model: add NVIDIA Nemotron-3-Puzzle-75B-A9B (NemotronHPuzzle) support by YanissAmz · Pull Request #25444 · ggml-org/llama.cpp
2.70T2 sourcer/LocalLLaMA
r/LocalLLaMAoriginal source ↗
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1w60jr5/model_add_nvidia_nemotron3puzzle75ba9b/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryPull request to llama.cpp adds inference support for NVIDIA Nemotron-3-Puzzle-75B-A9B, a mixture-of-experts model, via the ggml backend.
Why it mattersTracks which architectures are now runnable locally; useful for users selecting models within the llama.cpp stack.
Cited by
No citations on record.
