Skip to content

Open systems / 03

Open-source releases worth evaluating.

Find new AI tools and meaningful releases with source links, adoption context, and practical tradeoffs.

  1. ggml-org/llama.cpp b10730

    <details open> qwen4exp: sum the indexer heads by slices (#28023) * qwen4exp: sum the indexer heads by slices The head reduction went through a transpose and a sum_rows over ne[1], which left sum_rows with ne0 = 4, one block per row for a four element reduction, and the transpos…

    github.com17 days agoView details

  2. ggml-org/llama.cpp b10729

    <details open> metal : add fa-vec tunings for M1 Ultra (#28088) * metal : add fa-vec tunings for M1 Ultra * metal : move M1 Ultra tunings after M1 Max section * metal : remove duplicate blank line </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.…

    github.com17 days agoView details

  3. nexu-io/html-anything

    ✨ The agentic HTML editor — your local AI agent writes the HTML, you ship it. 🚀 75 Skills × 9 Surfaces (magazine · deck · poster · XHS / tweet · prototype · data report · Hyperframes) 🛡️ Sandboxed preview · 📤 1-click to WeChat / X / Zhihu / HTML / PNG 🔑 Zero API key — Claude…

    github.com4 months ago39 ptsView details

  4. modelcontextprotocol/servers 2026.8.31

    # Release : v2026.8.31 ## Updated packages - @modelcontextprotocol/server-filesystem@2026.8.31 - @modelcontextprotocol/server-memory@2026.8.31 - @modelcontextprotocol/server-sequential-thinking@2026.8.31 - @modelcontextprotocol/server-everything@2026.8.31

    github.com17 days agoView details

  5. Jwuthri/Tracely-ai

    Trace-native CI/CD for AI agents — production failures become regression tests that block the PR. Auto-detect, cluster, freeze into hermetic cases, replay in CI for $0.

    github.com3 months ago31 ptsView details

  6. MontyGovernance/montycat-mcp

    Shared, persistent memory for AI agents. Self-hosted MCP server with semantic search, vector RAG, and live updates. Works with Claude, Cursor, Codex, and any MCP client.

    github.com1 month ago17 ptsView details

  7. nexu-io/open-design

    🎨 The open-source Claude Design alternative. 🖥️ Local-first desktop app. 🖼️ Your coding agent becomes the design engine: prototypes, landing pages, dashboards, slides, images & video — real files, HTML/PDF/PPTX/MP4 export. 🤖 Claude Code / Codex / Cursor / Gemini / OpenCode /…

    github.com4 months ago49 ptsView details

  8. DEEIX-AI/DEEIX-Chat

    An enterprise AI workspace for model routing, multimodal chat, files, tools, billing, identity, and operations.

    github.com4 months ago32 ptsView details

  9. ggml-org/llama.cpp b10712

    <details open> vulkan: top_k radix select for k >= 1024 for Qwen 3.8 Flash Next (#28032) * vulkan: add top-k radix sort shader for k >= 1024 * add Qwen 3.8 Flash Next top-k tests * add top-k qsa fusion * clean up code </details> **Website:** - <https://llama.app> **Attestations:…

    github.com18 days agoView details

  10. ggml-org/llama.cpp b10711

    <details open> hexagon: fix CPY fence bug (#28033) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/44046740> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/downlo…

    github.com18 days agoView details

  11. ggml-org/llama.cpp b10710

    <details open> metal : add remaining Q4_1/Q5_0/Q5_1 fa-vec tunings for M2 (#28017) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/44044108> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/gg…

    github.com18 days agoView details

  12. DenisSergeevitch/agents-best-practices

    Provider-neutral Agent Skill for Codex, Claude Code, and agentic harness design.

    github.com4 months ago34 ptsView details

  13. opensquilla/opensquilla

    OpenSquilla — Token-Efficient AI Agent with same budget, higher intelligence density

    github.com4 months ago38 ptsView details

  14. codecoradev/uteke

    🧠 The Brain for Your AI — Local-first memory engine for AI agents. Store, recall, and search memories with semantic embeddings. Single Rust binary, zero config, fully offline.

    github.com3 months ago24 ptsView details

  15. tmasjc/deka-oss

    Deka — Aligning Human Intuition with Semantic Space

    github.com3 months ago17 ptsView details

  16. t8y2/dbx

    20 MB lightweight cross-platform database client for 90+ databases, including MySQL, PostgreSQL, SQLite, Redis, MongoDB, DuckDB, SQL Server, and Dameng. Built-in AI, MCP Server, CLI, desktop and Docker. | 轻量级跨平台数据库管理工具,支持 MySQL、PostgreSQL、SQLite、Redis、MongoDB、达梦等 90+ 数据库,提供桌面端、D…

    github.com4 months ago43 ptsView details

  17. ggml-org/llama.cpp b10690

    <details open> memory : copy Hadamard matrix to k_rot tensor only if it has buffer assigned to prevent crashes during context shift of unquantized K cache (#27967) Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com> Co-authored-by: AesSedai <7980540+AesSedai@users.noreply.gi…

    github.com19 days agoView details

  18. ggml-org/llama.cpp b10689

    <details open> ggml: allow passing alloc dependencies in graph_optimize (#27301) * ggml: allow passing alloc dependencies in graph_optimize * add alloc dep tests * add TODO about using flat array </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.c…

    github.com19 days agoView details

  19. WenyuChiou/awesome-agentic-ai-zh

    A trilingual (繁中 / English / 简中) learning roadmap for agentic AI: from LLM basics to multi-agent systems, with 240+ curated resources and hands-on examples. 中文 AI agent 學習地圖。

    github.com4 months ago38 ptsView details

  20. ggml-org/llama.cpp b10688

    <details open> metal : add fa-vec tunings for M2 (#27940) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/43902029> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases…

    github.com19 days agoView details

  21. EthanYoQ/AI-Novel-Writer

    AI 小说创作软件:把灵感、角色、世界观、大纲、章节写作、审稿和修稿组织成可控流程;提供 Windows/macOS 桌面版,支持本地和在线模型。AI Novel Writing Software: Organizes inspirations, characters, worldbuilding, outlines, chapter drafting, review, and revision into a controllable workflow. Features desktop apps for Windows/macOS, Ollama i…

    github.com2 months ago30 ptsView details

  22. virgiliojr94/book-to-skill

    Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.

    github.com4 months ago45 ptsView details

  23. esengine/DeepSeek-Reasonix

    DeepSeek-native AI coding agent for your terminal. Engineered around prefix-cache stability — leave it running.

    github.com5 months ago46 ptsView details

  24. kldhsh123/Afterglow

    使用社交软件聊天记录结合向量数据库让AI更好的扮演对方的角色,在不微调模型的情况下可以达到可观的效果。把曾经的美好,续成往后的陪伴。

    github.com3 months ago25 ptsView details

  25. juanjuandog/FinSight-AI

    AI equity research agent with resilient workflows, evidence-grounded RAG, versioned reports, and automated quality evaluation.

    github.com4 months ago30 ptsView details

  26. Deuz-AI/Deuz-SDK

    Zero-dependency TypeScript framework for production AI agents: durable execution, long-term memory, hybrid RAG, MCP tool calling, human-in-the-loop approval, planning and CodeAct sandboxes. One streaming API for Claude, GPT, Gemini, Grok, Mistral and DeepSeek — Node, Bun, Deno,…

    github.com3 months ago31 ptsView details

  27. freestylefly/awesome-gpt-image-2

    Prompt as Code | GPT Image 2 / 2.5 提示词与案例库,530+ 个案例、20+ 套工业级模板与可复用 Skills,新增 2.5 同提示词对比专区,附完整提示词与生成记录,持续更新。

    github.com4 months ago45 ptsView details

  28. GetBusbar/busbar

    The execution control plane for AI agents. Govern every model request, MCP tool call, A2A delegation, and downstream action—while keeping your native protocols intact: six native LLM protocols in, six native LLM protocols out.

    github.com3 months ago22 ptsView details

  29. openai/openai-python v3.6.0

    ## [3.6.0](https://github.com/openai/openai-python/compare/v3.5.0...v3.6.0) (2026-08-27) ### Features * **api:** add compute_units to Responses and Chat Completions usage ([#3749](https://github.com/openai/openai-python/issues/3749)) ([52421d1](https://github.com/openai/openai-p…

    github.com20 days agoView details

  30. volcengine/SearchCLI

    Open CLI for integrating AI search, recommendation, and conversational retrieval into agent systems and business systems

    github.com4 months ago31 ptsView details

  31. stuinfla/Repo-Explainer

    Repo Explainer — turn any GitHub repo into a visual explainer page. Pipeline + 5 live examples.

    github.com3 months ago18 ptsView details

  32. ggml-org/llama.cpp b10679

    <details open> bench: add --tensor-read-lazy (#27881) * bench: add --tensor-read-lazy * rm the alias * rename to LLAMA_LAZY_MODE_* </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/43739963> **macOS/iOS:** - [ma…

    github.com20 days agoView details

  33. ggml-org/llama.cpp b10678

    <details open> model: qwen4exp: reduce number of graph splits (#27880) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/43734155> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama…

    github.com20 days agoView details

  34. earthtojake/text-to-cad

    A library of agent skills for CAD, CAE and CAM

    github.com4 months ago42 ptsView details

  35. ggml-org/llama.cpp b10677

    <details open> vulkan: fix missing view-alias dependencies in ggml_vk_graph_optimize (#27812) * vulkan: fix missing view-alias dependencies in ggml_vk_graph_optimize is_src_of doesn't treat two views of one tensor as dependent, so the optimizer reorders nodes across aliased read…

    github.com20 days agoView details

  36. langchain-ai/langchain langchain==1.4.0a2

    Alpha preview of `langchain.mcp` — a first-party adapter that turns any MCP server into LangChain tools you can hand straight to `create_agent`. Connection handling is [FastMCP](https://gofastmcp.com/clients/client)'s, so its client features are available as-is rather than re-im…

    github.com20 days agoView details

  37. ruvnet/agenticow

    Git for Agent Memory: Copy-On-Write vector branching for embedded multi-agent memory (83x faster, 3000x smaller snapshots)

    github.com2 months ago17 ptsView details

  38. crynta/terax-ai

    Lightweight (7MB) Terminal-first AI-native dev workspace

    github.com4 months ago40 ptsView details

  39. ongridio/ongrid

    An ops AI Agent that understands your infrastructure, finds the root cause, and fixes it — right from Slack, Telegram, Lark or DingTalk.

    github.com3 months ago30 ptsView details

  40. ConardLi/garden-skills

    ConardLi's open-source Skills collection, featuring web design, knowledge retrieval, image generation, and more.

    github.com5 months ago41 ptsView details

  41. ggml-org/llama.cpp b10666

    <details open> tests : run test-save-load-state across all architectures (#27755) * tests : run test-save-load-state across all architectures test-save-load-state previously only ran in ctest against a single downloaded model (tinyllamas/stories15M), i.e. only the llama arch. Ad…

    github.com21 days agoView details

  42. Ontos-AI/knowhere

    Knowhere extracts, parses, and outputs structured chunks ready for AI Agents and RAG.

    github.com4 months ago35 ptsView details

  43. browser-use/browser-harness

    Browser Harness | Self-healing harness that enables LLMs to complete any task.

    github.com5 months ago42 ptsView details

  44. ggml-org/llama.cpp b10665

    <details open> model: add DSpark support for Nemotron3.5 (#27804) * model: add DSpark support for Nemotron3.5 * Update src/models/dflash.cpp Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co> --------- Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingf…

    github.com21 days agoView details

  45. VigoZhao/AI-Visual-Prompt-Cookbook

    118+ plug-and-play JSON style packs for Nano Banana Pro, GPT Image & Midjourney. Copy one JSON, get a style. Updated daily.

    github.com4 months ago28 ptsView details

  46. JuliusBrussee/caveman

    🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking like a caveman.

    github.com5 months ago50 ptsView details

  47. ggml-org/llama.cpp b10664

    <details open> ggml-hexagon: add HTP unary ops for ABS and LOG (#27786) Add HVX-accelerated implementations for GGML_OP_LOG and GGML_UNARY_OP_ABS on the HTP backend. - Register HTP_OP_UNARY_ABS and HTP_OP_UNARY_LOG in op_remap_to_htp() - Add ABS and LOG to ggml_backend_hexagon_d…

    github.com21 days agoView details

  48. langchain-ai/langchain langchain==1.4.0a1

    Initial release fix(langchain): name the content type MCP conversion could not handle release(langchain): 1.4.0a1 test(langchain): skip MCP tests on a pydantic older than `mcp` supports test(langchain): drive MCP tests through FastMCP's own utilities fix(langchain/mcp): review e…

    github.com21 days agoView details

  49. langchain-ai/langchain langchain-fireworks==1.6.1

    Changes since langchain-fireworks==1.6.0 release(fireworks): 1.6.1 (#39975) fix(fireworks): drop reasoning history blocks (#39973) chore(model-profiles): refresh model profile data (#39844)

    github.com21 days agoView details

  50. shyftlabs/continuum

    Continuum — the agent runtime by ShyftLabs. Build, orchestrate, ship.

    github.com3 months ago19 ptsView details

Newsletter

Get practical AI engineering notes

Receive source-checked analysis of models, agents, evaluation, retrieval, and production reliability. Sent only when there is useful work to share.