Open systems / 03
Open-source releases worth evaluating.
Find new AI tools and meaningful releases with source links, adoption context, and practical tradeoffs.
<details open> qwen4exp: sum the indexer heads by slices (#28023) * qwen4exp: sum the indexer heads by slices The head reduction went through a transpose and a sum_rows over ne[1], which left sum_rows with ne0 = 4, one block per row for a four element reduction, and the transpos…
github.com17 days agoView details
<details open> metal : add fa-vec tunings for M1 Ultra (#28088) * metal : add fa-vec tunings for M1 Ultra * metal : move M1 Ultra tunings after M1 Max section * metal : remove duplicate blank line </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.…
github.com17 days agoView details
✨ The agentic HTML editor — your local AI agent writes the HTML, you ship it. 🚀 75 Skills × 9 Surfaces (magazine · deck · poster · XHS / tweet · prototype · data report · Hyperframes) 🛡️ Sandboxed preview · 📤 1-click to WeChat / X / Zhihu / HTML / PNG 🔑 Zero API key — Claude…
github.com4 months ago39 ptsView details
modelcontextprotocol/servers 2026.8.31
# Release : v2026.8.31 ## Updated packages - @modelcontextprotocol/server-filesystem@2026.8.31 - @modelcontextprotocol/server-memory@2026.8.31 - @modelcontextprotocol/server-sequential-thinking@2026.8.31 - @modelcontextprotocol/server-everything@2026.8.31
github.com17 days agoView details
Trace-native CI/CD for AI agents — production failures become regression tests that block the PR. Auto-detect, cluster, freeze into hermetic cases, replay in CI for $0.
github.com3 months ago31 ptsView details
Shared, persistent memory for AI agents. Self-hosted MCP server with semantic search, vector RAG, and live updates. Works with Claude, Cursor, Codex, and any MCP client.
github.com1 month ago17 ptsView details
🎨 The open-source Claude Design alternative. 🖥️ Local-first desktop app. 🖼️ Your coding agent becomes the design engine: prototypes, landing pages, dashboards, slides, images & video — real files, HTML/PDF/PPTX/MP4 export. 🤖 Claude Code / Codex / Cursor / Gemini / OpenCode /…
github.com4 months ago49 ptsView details
An enterprise AI workspace for model routing, multimodal chat, files, tools, billing, identity, and operations.
github.com4 months ago32 ptsView details
<details open> vulkan: top_k radix select for k >= 1024 for Qwen 3.8 Flash Next (#28032) * vulkan: add top-k radix sort shader for k >= 1024 * add Qwen 3.8 Flash Next top-k tests * add top-k qsa fusion * clean up code </details> **Website:** - <https://llama.app> **Attestations:…
github.com18 days agoView details
<details open> hexagon: fix CPY fence bug (#28033) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/44046740> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/downlo…
github.com18 days agoView details
<details open> metal : add remaining Q4_1/Q5_0/Q5_1 fa-vec tunings for M2 (#28017) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/44044108> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/gg…
github.com18 days agoView details
DenisSergeevitch/agents-best-practices
Provider-neutral Agent Skill for Codex, Claude Code, and agentic harness design.
github.com4 months ago34 ptsView details
OpenSquilla — Token-Efficient AI Agent with same budget, higher intelligence density
github.com4 months ago38 ptsView details
🧠 The Brain for Your AI — Local-first memory engine for AI agents. Store, recall, and search memories with semantic embeddings. Single Rust binary, zero config, fully offline.
github.com3 months ago24 ptsView details
Deka — Aligning Human Intuition with Semantic Space
github.com3 months ago17 ptsView details
20 MB lightweight cross-platform database client for 90+ databases, including MySQL, PostgreSQL, SQLite, Redis, MongoDB, DuckDB, SQL Server, and Dameng. Built-in AI, MCP Server, CLI, desktop and Docker. | 轻量级跨平台数据库管理工具,支持 MySQL、PostgreSQL、SQLite、Redis、MongoDB、达梦等 90+ 数据库,提供桌面端、D…
github.com4 months ago43 ptsView details
<details open> memory : copy Hadamard matrix to k_rot tensor only if it has buffer assigned to prevent crashes during context shift of unquantized K cache (#27967) Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com> Co-authored-by: AesSedai <7980540+AesSedai@users.noreply.gi…
github.com19 days agoView details
<details open> ggml: allow passing alloc dependencies in graph_optimize (#27301) * ggml: allow passing alloc dependencies in graph_optimize * add alloc dep tests * add TODO about using flat array </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.c…
github.com19 days agoView details
WenyuChiou/awesome-agentic-ai-zh
A trilingual (繁中 / English / 简中) learning roadmap for agentic AI: from LLM basics to multi-agent systems, with 240+ curated resources and hands-on examples. 中文 AI agent 學習地圖。
github.com4 months ago38 ptsView details
<details open> metal : add fa-vec tunings for M2 (#27940) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/43902029> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases…
github.com19 days agoView details
AI 小说创作软件:把灵感、角色、世界观、大纲、章节写作、审稿和修稿组织成可控流程;提供 Windows/macOS 桌面版,支持本地和在线模型。AI Novel Writing Software: Organizes inspirations, characters, worldbuilding, outlines, chapter drafting, review, and revision into a controllable workflow. Features desktop apps for Windows/macOS, Ollama i…
github.com2 months ago30 ptsView details
Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.
github.com4 months ago45 ptsView details
DeepSeek-native AI coding agent for your terminal. Engineered around prefix-cache stability — leave it running.
github.com5 months ago46 ptsView details
使用社交软件聊天记录结合向量数据库让AI更好的扮演对方的角色,在不微调模型的情况下可以达到可观的效果。把曾经的美好,续成往后的陪伴。
github.com3 months ago25 ptsView details
AI equity research agent with resilient workflows, evidence-grounded RAG, versioned reports, and automated quality evaluation.
github.com4 months ago30 ptsView details
Zero-dependency TypeScript framework for production AI agents: durable execution, long-term memory, hybrid RAG, MCP tool calling, human-in-the-loop approval, planning and CodeAct sandboxes. One streaming API for Claude, GPT, Gemini, Grok, Mistral and DeepSeek — Node, Bun, Deno,…
github.com3 months ago31 ptsView details
freestylefly/awesome-gpt-image-2
Prompt as Code | GPT Image 2 / 2.5 提示词与案例库,530+ 个案例、20+ 套工业级模板与可复用 Skills,新增 2.5 同提示词对比专区,附完整提示词与生成记录,持续更新。
github.com4 months ago45 ptsView details
The execution control plane for AI agents. Govern every model request, MCP tool call, A2A delegation, and downstream action—while keeping your native protocols intact: six native LLM protocols in, six native LLM protocols out.
github.com3 months ago22 ptsView details
## [3.6.0](https://github.com/openai/openai-python/compare/v3.5.0...v3.6.0) (2026-08-27) ### Features * **api:** add compute_units to Responses and Chat Completions usage ([#3749](https://github.com/openai/openai-python/issues/3749)) ([52421d1](https://github.com/openai/openai-p…
github.com20 days agoView details
Open CLI for integrating AI search, recommendation, and conversational retrieval into agent systems and business systems
github.com4 months ago31 ptsView details
Repo Explainer — turn any GitHub repo into a visual explainer page. Pipeline + 5 live examples.
github.com3 months ago18 ptsView details
<details open> bench: add --tensor-read-lazy (#27881) * bench: add --tensor-read-lazy * rm the alias * rename to LLAMA_LAZY_MODE_* </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/43739963> **macOS/iOS:** - [ma…
github.com20 days agoView details
<details open> model: qwen4exp: reduce number of graph splits (#27880) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/43734155> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama…
github.com20 days agoView details
A library of agent skills for CAD, CAE and CAM
github.com4 months ago42 ptsView details
<details open> vulkan: fix missing view-alias dependencies in ggml_vk_graph_optimize (#27812) * vulkan: fix missing view-alias dependencies in ggml_vk_graph_optimize is_src_of doesn't treat two views of one tensor as dependent, so the optimizer reorders nodes across aliased read…
github.com20 days agoView details
langchain-ai/langchain langchain==1.4.0a2
Alpha preview of `langchain.mcp` — a first-party adapter that turns any MCP server into LangChain tools you can hand straight to `create_agent`. Connection handling is [FastMCP](https://gofastmcp.com/clients/client)'s, so its client features are available as-is rather than re-im…
github.com20 days agoView details
Git for Agent Memory: Copy-On-Write vector branching for embedded multi-agent memory (83x faster, 3000x smaller snapshots)
github.com2 months ago17 ptsView details
Lightweight (7MB) Terminal-first AI-native dev workspace
github.com4 months ago40 ptsView details
An ops AI Agent that understands your infrastructure, finds the root cause, and fixes it — right from Slack, Telegram, Lark or DingTalk.
github.com3 months ago30 ptsView details
ConardLi's open-source Skills collection, featuring web design, knowledge retrieval, image generation, and more.
github.com5 months ago41 ptsView details
<details open> tests : run test-save-load-state across all architectures (#27755) * tests : run test-save-load-state across all architectures test-save-load-state previously only ran in ctest against a single downloaded model (tinyllamas/stories15M), i.e. only the llama arch. Ad…
github.com21 days agoView details
Knowhere extracts, parses, and outputs structured chunks ready for AI Agents and RAG.
github.com4 months ago35 ptsView details
Browser Harness | Self-healing harness that enables LLMs to complete any task.
github.com5 months ago42 ptsView details
<details open> model: add DSpark support for Nemotron3.5 (#27804) * model: add DSpark support for Nemotron3.5 * Update src/models/dflash.cpp Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co> --------- Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingf…
github.com21 days agoView details
VigoZhao/AI-Visual-Prompt-Cookbook
118+ plug-and-play JSON style packs for Nano Banana Pro, GPT Image & Midjourney. Copy one JSON, get a style. Updated daily.
github.com4 months ago28 ptsView details
🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking like a caveman.
github.com5 months ago50 ptsView details
<details open> ggml-hexagon: add HTP unary ops for ABS and LOG (#27786) Add HVX-accelerated implementations for GGML_OP_LOG and GGML_UNARY_OP_ABS on the HTP backend. - Register HTP_OP_UNARY_ABS and HTP_OP_UNARY_LOG in op_remap_to_htp() - Add ABS and LOG to ggml_backend_hexagon_d…
github.com21 days agoView details
langchain-ai/langchain langchain==1.4.0a1
Initial release fix(langchain): name the content type MCP conversion could not handle release(langchain): 1.4.0a1 test(langchain): skip MCP tests on a pydantic older than `mcp` supports test(langchain): drive MCP tests through FastMCP's own utilities fix(langchain/mcp): review e…
github.com21 days agoView details
langchain-ai/langchain langchain-fireworks==1.6.1
Changes since langchain-fireworks==1.6.0 release(fireworks): 1.6.1 (#39975) fix(fireworks): drop reasoning history blocks (#39973) chore(model-profiles): refresh model profile data (#39844)
github.com21 days agoView details
Continuum — the agent runtime by ShyftLabs. Build, orchestrate, ship.
github.com3 months ago19 ptsView details
Newsletter
Get practical AI engineering notes
Receive source-checked analysis of models, agents, evaluation, retrieval, and production reliability. Sent only when there is useful work to share.