Open source / #20535
Open source
ggml-org/llama.cpp b10776
GitHub Releases · github-actions[bot] · 15 days ago
<details open> model: add NVIDIA Nemotron-3-Puzzle-75B-A9B (NemotronHPuzzle) support (#25444) * hparams: add per-layer n_ff_exp/n_expert_used arrays with scalar-or-array loading G1/G2 infrastructure for variable-per-layer expert FFN size and top-k routing (required for Puzzle-75…
Source and ranking details
- Source adapter
- github-releases
- Source weight
- 3
- Points
- 0
- Age
- 362.1 h
- Stored score
- 0.00005
- Current score
- 0.00005
- First seen
- 2026-09-03T08:37:05.988Z
Newsletter
Get practical AI engineering notes
Receive source-checked analysis of models, agents, evaluation, retrieval, and production reliability. Sent only when there is useful work to share.