Skip to content

Open source / #20535

Open source

ggml-org/llama.cpp b10776

GitHub Releases · github-actions[bot] · 15 days ago

<details open> model: add NVIDIA Nemotron-3-Puzzle-75B-A9B (NemotronHPuzzle) support (#25444) * hparams: add per-layer n_ff_exp/n_expert_used arrays with scalar-or-array loading G1/G2 infrastructure for variable-per-layer expert FFN size and top-k routing (required for Puzzle-75…

Source adapter
github-releases
Source weight
3
Points
0
Age
362.1 h
Stored score
0.00005
Current score
0.00005
First seen
2026-09-03T08:37:05.988Z

Newsletter

Get practical AI engineering notes

Receive source-checked analysis of models, agents, evaluation, retrieval, and production reliability. Sent only when there is useful work to share.