Open systems / 03
Open-source releases worth evaluating.
Find new AI tools and meaningful releases with source links, adoption context, and practical tradeoffs.
## [3.11.0](https://github.com/openai/openai-python/compare/v3.10.0...v3.11.0) (2026-09-09) ### Features * **api:** Add expiration controls for service account keys ([#3825](https://github.com/openai/openai-python/issues/3825)) ([f348ec8](https://github.com/openai/openai-python/…
github.com8 days agoView details
# v0.29.0 ## Highlights This release features 594 commits from 277 contributors (91 new)! * **Model Runner V2 is now the default for all models** (#53183), completing the rollout that began with pooling models (#48290). MRV2 also gained CUDA graph memory profiling for KV cache a…
github.com9 days agoView details
A 2.78-trillion-parameter Kimi K3 running inference on a single CPU in 8.24 GB of RAM. Portable C99: no BLAS, no framework, no GPU.
github.com1 month ago39 ptsView details
<details open> jinja: treat a null left operand of in as a plain lookup (#28620) Templates that default an optional variable to none and then test its membership in a map hit an error, while the same expression is a normal lookup returning false in Jinja. The undefined counterpa…
github.com9 days agoView details
<details open> vulkan: add dedicated iq4_xs mat-vec shader (#28426) * vulkan: add dedicated iq4_xs mat-vec shader Dedicated mul_mat_vec_iq4_xs for the dmmv path, replacing the generic fallback. ~+6-17% token generation on RDNA4 depending on model. Assisted-by: Pi agent with Qwen…
github.com9 days agoView details
<details open> vulkan: add f16 B-type matmul pipelines and warp tile size tuning for Intel coopmat1 (#27471) * vulkan: add f16 B-type matmul pipelines and warp tile size tuning for Intel coopmat1 * simplify mmp selection in mul_mat_id per review comment * vulkan: enable f16 B-ty…
github.com9 days agoView details
## [3.10.0](https://github.com/openai/openai-python/compare/v3.9.0...v3.10.0) (2026-09-08) ### Features * **api:** add GPT Image 2.5 models and image options ([#3824](https://github.com/openai/openai-python/issues/3824)) ([5b39c45](https://github.com/openai/openai-python/commit/…
github.com9 days agoView details
Clone any viral video with AI agents. Not just a script, the whole workflow: swap the face, the words, the B-roll, ship 100 variants in one command, and get your 100M views.
github.com1 month ago40 ptsView details
## [3.9.0](https://github.com/openai/openai-python/compare/v3.8.0...v3.9.0) (2026-09-05) ### Features * **api:** Add prompt cache diagnostics ([#3800](https://github.com/openai/openai-python/issues/3800)) ([8326784](https://github.com/openai/openai-python/commit/83267847a0219ea8…
github.com9 days agoView details
Evolutionary multi-agent runtime that breeds, evaluates, and improves autonomous agents across reproducible epochs to converge on optimization of a goal.
github.com1 month ago25 ptsView details
langchain-ai/langchain langchain-openai==1.6.1
Changes since langchain-openai==1.6.0 fix(openai): bump `max_completion_tokens` in cache breakpoint integration test (#40284) release(openai): 1.6.1 (#40268) chore(model-profiles): refresh model profile data (#40217) fix(openai): support Azure AD auth with OpenAI 3.8 (#40190) fe…
github.com9 days agoView details
Use ChatGPT Web (including Pro) as a native model in Codex — with context, tools, streaming and images, without using Codex quota.
github.com1 month ago40 ptsView details
Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
github.com3 months ago52 ptsView details
<details open> model : support Kimi-K3 recurrent-state rollback (#28466) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/45866131> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/lla…
github.com10 days agoView details
<details open> hexagon: add RELU and LEAKY_RELU ops (#28585) * hexagon: add RELU op * hexagon: add LEAKY_RELU op too </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/45844223> **macOS/iOS:** - [macOS Apple Sili…
github.com10 days agoView details
🪐 HelixWorld: real-time interactive audio-visual world model.
github.com1 month ago28 ptsView details
<details open> tests : initialize the L2_NORM batch array (#28553) * tests: bind the L2_NORM batch count to a local GCC cannot prove the loop fills norms up to the index read after it while the bound is a class member, so it reports a maybe uninitialized use. Reading the count o…
github.com10 days agoView details
The open-source agent harness - the runtime layer that turns an LLM into a working agent.
github.com1 month ago38 ptsView details
Cheaper and better Greptile alternative runs on your own github actions.
github.com1 month ago22 ptsView details
<details open> caps : recheck typed content if template checks for string (#28511) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/45690927> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/gg…
github.com11 days agoView details
Universal provider proxy for OpenAI Codex & Claude Code — use any LLM (Claude, Gemini, Grok, DeepSeek, Ollama…) with Codex CLI, App, SDK, and Claude Code
github.com3 months ago42 ptsView details
<details open> ggml-cuda: fix divergent barrier in f16 flash attention (#27870) * ggml-cuda: fix divergent barrier in f16 flash attention * ggml-cuda: avoid duplicate metadata pointer setup </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggm…
github.com11 days agoView details
<details open> ggml: allow backend inputs to not create another split (#28387) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/45674860> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-o…
github.com11 days agoView details
davidahmann/applied-ai-field-guide
The Applied AI Field Guide: fieldwork, value, engineering, and operations for AI that works beyond the demo.
github.com1 month ago20 ptsView details
Graph-Orchestrated Agent Loop — a production-grade framework on LangGraph. Combine workflow graphs and agent loops, transpile Dify DSL to runnable code, swap wire protocols (Dify/OpenAI).
github.com1 month ago21 ptsView details
autonomous red teaming platform; multi-agent offensive-security meta-harness
github.com2 months ago38 ptsView details
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
github.com2 months ago38 ptsView details
Open-source infrastructure for the full inference lifecycle: deploy, observe, scale, optimize, and safely release self-hosted models behind one endpoint.
github.com1 month ago18 ptsView details
Turbocharge Claude Code, Cursor, Codex, Gemini & every coding agent: faster, cheaper, with contextual understanding specific to your codebase.
github.com2 months ago37 ptsView details
Turbocharge Claude Code, Cursor, Codex, Gemini & every coding agent: faster, cheaper, with contextual understanding specific to your codebase.
github.com2 months ago39 ptsView details
Open-source auth gateway connecting 1500+ SaaS providers to AI agents through SDK, CLI, MCP, HTTP, and OpenAPI.
github.com2 months ago38 ptsView details
Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.
github.com3 months ago40 ptsView details
🔥 On-Policy Self-Distillation in Diffusion Models
github.com1 month ago27 ptsView details
A persistent workspace for development work that self-improves and continues beyond one session.
github.com2 months ago36 ptsView details
FDE and AI engineering guide for production systems: value, architecture, evals, security, deployment, and operations.
github.com1 month ago18 ptsView details
cobusgreyling/loop-engineering
Practical patterns, starters & CLI tools for loop engineering with AI coding agents. Design systems that prompt and orchestrate agents (inspired by Addy Osmani and Boris Cherny). Includes loop-audit, loop-init, loop-cost.
github.com3 months ago41 ptsView details
Drop-in AI memory layer with 2x faster retrieval and 10x lower cost. Fully compatible with Mem0 API. Migrate in 5 minutes without any code changes. Self-host for free.
github.com1 month ago21 ptsView details
See what your coding agents did and what it cost. Breaks each task down into work steps — tools used, files changed, tests run, time and tokens spent. Local-first dashboard for Claude Code, Codex, OpenCode, and more. No login, no telemetry.
github.com1 month ago29 ptsView details
Turn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas Cloud + ffmpeg. An agent skill.
github.com2 months ago33 ptsView details
XERJ is the new way for AI to search data. Its autoindex capability activates agents to know your data without the token waste of grep and sed. One command indexes code, docs, logs and PDFs for search, RAG, security audits and agent memory, using 40x fewer tokens than grep. Elas…
github.com2 months ago33 ptsView details
Waku Waku! Waku Agent is a local-first AI agent harness you actually own, including loop, memory, eval, all in code built to stay legible as it grows.
github.com2 months ago33 ptsView details
AI logo animation skill: turn raster logos into smooth SVG animation, animated HTML demos, GIF/video previews, and motion QA evidence.
github.com3 months ago34 ptsView details
OpenRouter for agent tools. Join community here: https://discord.gg/6mQYYfFMAn
github.com2 months ago32 ptsView details
All-in-One Multimodal Parsing Engine + Ontology-Powered, LLM Wiki-Driven AI-Ready Knowledge Engine
github.com2 months ago30 ptsView details
Flatkey media generation CLI for images, videos, audio, text, credits, and model discovery.
github.com2 months ago30 ptsView details
<details open> metal : fix memory leak in early return (#28399) </details> **Website:** - <https://llama.app> **Attestations:** - <https://github.com/ggml-org/llama.cpp/attestations/45438612> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/re…
github.com13 days agoView details
<details open> sycl : fix test-backend-ops CI break && restore Kronecker product FWHT support (#28016) (#28254) * Reapply "sycl : add Kronecker product FWHT support for sizes 384, 640, 768, 12…" (#28184) This reverts commit c845263f8b7d60113e213a3bd2d5cc6472ccf204. * tests : fix…
github.com13 days agoView details
<details open> sycl: attribute device allocations by site (GGML_SYCL_MEMTRACE) (#27631) define two new environment variables to better understand how much memory is being allocated, and when. This has been invaluable in inproving the --fit algorithm, and is likely to be useful w…
github.com13 days agoView details
Open-source AI research workbench for scientific research—local-first and model-agnostic.
github.com2 months ago37 ptsView details
Encrypted, fully offline agentic memory. One click install, GUI w/ memory map, all OS and agents. Superior memory creation, storage and retrieval.
github.com1 month ago29 ptsView details
Newsletter
Get practical AI engineering notes
Receive source-checked analysis of models, agents, evaluation, retrieval, and production reliability. Sent only when there is useful work to share.