Skip to content

Open source / #23728

Open source

ggml-org/llama.cpp b10835

GitHub Releases · github-actions[bot] · 11 days ago

<details open> ggml-cuda: fix divergent barrier in f16 flash attention (#27870) * ggml-cuda: fix divergent barrier in f16 flash attention * ggml-cuda: avoid duplicate metadata pointer setup </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggm…

Source adapter
github-releases
Source weight
3
Points
0
Age
265.7 h
Stored score
0.00009
Current score
0.00009
First seen
2026-09-07T08:37:05.812Z

Newsletter

Get practical AI engineering notes

Receive source-checked analysis of models, agents, evaluation, retrieval, and production reliability. Sent only when there is useful work to share.