Release Calendar

Showing 81 releases in August 2026

Fri, Aug 28

(3)
paid
Exa
Dynamic Highlights

Dynamic Highlights is an advanced search primitive that treats tokens, rather than documents, as the fundamental unit of information retrieval. By analyzing all retrieved documents in a single forward pass, it dynamically allocates token budgets to the most relevant content, significantly reducing context bloat for AI agents.

open
Tencent
Hy4

Tencent Hy4 preview is a high-performance, open-source large language model designed for real-world productivity and complex reasoning tasks. It represents the latest iteration in the Tencent Hunyuan (Hy) series, focusing on enhanced agent capabilities and seamless integration into professional workflows.

SmolVM

SmolVM is an open-source AI sandbox infrastructure that provides disposable, hardware-isolated microVMs for AI agents. It features a unified API for VMMs like Firecracker, QEMU, and libkrun, allowing agents to safely execute code, browse the web, and manage persistent state.

Thu, Aug 27

(2)
paid
Cohere
Parse

Cohere Parse is a high-throughput, 2.3B-parameter vision language model designed to convert complex enterprise documents, tables, and images into structured, machine-readable Markdown. It is optimized for high-volume production environments, offering a cost-effective solution for document indexing, RAG, and agentic retrieval workflows.

open
Vercel
vgpu

vgpu is a modular, TypeScript-first WebGPU library designed for high-performance rendering in browsers, headless Node.js environments, and automated test suites. It features typed WGSL shader imports, a unified GPU context API, and built-in support for coding agents to interact with documentation and examples.

Wed, Aug 26

(1)
open
pnpm
pnpm

pnpm 12.0 is a major stable release that features a complete rewrite of the package manager in Rust. It maintains full compatibility with pnpm 11 commands, flags, and lockfile formats while delivering significant performance improvements.

Tue, Aug 25

(2)
open
IBM
Granite 4.2

Granite 4.2 is a series of open-source language and speech models designed for agentic enterprise workflows. These models feature native step-by-step reasoning, advanced tool-calling capabilities, and optimized architectures for deployment across cloud, on-premises, and edge environments.

open
Vercel
Run SDK

Run SDK is a secure, lightweight alternative to eval for executing untrusted JavaScript and TypeScript code within applications. It utilizes a hardened QuickJS sandbox to isolate code execution, allowing developers to expose specific host functions while preventing unauthorized access to system resources or secrets.

Mon, Aug 24

(4)
open
Microsoft
AutoSaddler

AutoSaddler is an open-source framework designed to automatically optimize LLM-agent harnesses by diagnosing execution traces. It improves agent performance by applying structured updates to prompts, tools, and middleware, ensuring that changes generalize across tasks rather than just fixing specific trajectories.

ElevenLabs CLI

ElevenLabs CLI v1 is a command-line interface that brings the entire ElevenLabs API into the terminal, specifically designed for developers and coding agents. It enables an 'agents-as-code' workflow, allowing users to manage voice agents through local configuration files, version control, and automated deployment pipelines.

GLiNER

GLiNER2.5 is a significant upgrade to the GLiNER architecture that replaces span enumeration with boundary prediction for more efficient, scalable information extraction. It introduces five key capabilities, including long-context processing, unlimited span length, joint entity-relation extraction, constrained classification, and span-level attributes.

free
AMD
Radeon Developer Panel (RDP)

The Radeon Developer Panel (RDP) is a specialized software tool designed for developers to optimize applications on AMD RDNA hardware. It serves as the primary interface for capturing RGP profiles, memory traces, raytracing scenes, and crash analysis dumps by communicating directly with the AMD Software: Adrenalin Edition driver.

Sat, Aug 22

(2)
Project SuperDex

Project SuperDex is a unified, open-source simulation platform designed for research into dexterous robotic manipulation. It integrates a contact-first physics engine, robotics authoring tools, and a scalable reinforcement learning interface to support complex hand-object interaction tasks.

open
EGOIST
Waku

Waku is a fast, native desktop application built in Rust and GPUI designed to unify various local coding agents into a single interface. It allows developers to manage projects, sessions, and transcripts locally while providing advanced features like Git-backed task rewinding and real-time agent steering.

Fri, Aug 21

(1)
free
Vercel
Is Agentic

Is Agentic is a diagnostic tool by Vercel that evaluates how effectively a website can be discovered, accessed, and utilized by AI agents. It provides a numeric readiness score along with actionable recommendations to improve compatibility with autonomous agent workflows.

Thu, Aug 20

(1)
open
OpenDCAI
3AGameFactory

3AGameFactory is a comprehensive open-source framework designed to automate the generation of 3A-quality game assets and engine-ready code. It utilizes a coding agent to orchestrate various pipelines for image, 3D object, motion, audio, and CG-video generation, supporting integration with Unreal Engine 5, Unity, Blender, and three.js.

Wed, Aug 19

(3)
open
RadixArk
Miles

Miles is an open-source reinforcement learning (RL) framework designed for large-scale model post-training, including agentic workloads. It integrates SGLang for high-throughput rollouts with Megatron-LM or FSDP2 for distributed training, providing a stable and efficient stack for frontier models.

Ornith-1.5

Ornith-1.5 is an open-source model family that advances foundation model development through an end-to-end self-improvement loop. The system autonomously proposes tasks, constructs task-specific scaffolds, and generates solution rollouts to continuously refine its reasoning, coding, and agentic capabilities.

S1-mini

S1-mini is a specialized 0.6B-parameter text normalization model designed to process raw speech-to-text transcripts. It cleans ASR output by removing fillers, resolving self-corrections, and applying proper punctuation, capitalization, and formatting for numbers, dates, and currency.

Tue, Aug 18

(3)
fx

fx is a tiny, high-performance coding agent harness and CLI tool written in Zig. It is designed for research and embeddability, offering a Unix-like experience that is model-agnostic and suitable for both local and cloud-based inference.

open
Lerd
Lerd

Lerd is an open-source, daemonless local PHP development environment for Linux and macOS. It utilizes rootless Podman containers to provide automatic .test domains, per-project PHP and Node isolation, and one-command TLS without requiring sudo or system pollution.

free
Meta
Meta XR Operator

Meta XR Operator is an experimental OpenXR API layer that enables MCP-compatible AI agents to visually perceive, navigate, and interact with running VR applications. It allows developers to automate the build-test-verify loop by letting agents perform tasks like UI navigation, scene inspection, and controller input simulation within the Meta XR Simulator.

Mon, Aug 17

(4)
open
FlashML
FreeToken

FreeToken is an edge-native Mixture-of-Experts (MoE) serving engine designed to run frontier-scale open-weight models on consumer hardware. It optimizes performance by treating heterogeneous edge resources like GPUs, CPUs, and memory as a unified, elastic inference platform.

Origin

Origin is a Git-compatible code hosting and collaboration platform designed specifically for the agentic era. It enables teams to host repositories, manage pull requests, and integrate seamlessly with AI agents to handle high-frequency code changes and automated workflows.

Qwen3.8-27B-Uncensored-MLX

Qwen3.8-27B-Uncensored-MLX is an abliterated, refusal-removed version of the Qwen3.8-27B vision-language model, specifically optimized for Apple Silicon using the MLX framework. It is designed for AI safety research, red-teaming, and interpretability studies, providing full access to the model's capabilities without built-in guardrails.

open
Tencent
UI-Mate

UI-Mate-27B is an open-weight foundation GUI agent designed for long-horizon desktop automation across various operating systems. It observes live screen states to reason and execute structured keyboard and mouse actions, supporting in-context demonstrations to improve task accuracy.

Sun, Aug 16

(2)
open
Alibaba
Qwen-UI-Agent

Qwen-UI-Agent is a real-world-centric foundation GUI agent designed to unify mobile, computer, browser, and DeepSearch scenarios into a single model. It enables autonomous task execution across platforms by combining GUI interaction with CLI execution, proactive service initiation, and long-horizon planning.

open
Google
SAM (Sovereign Agent Mesh)

SAM is a zero-config, zero-trust peer-to-peer network designed specifically for autonomous AI agents. It enables secure, environment-agnostic communication and tool sharing across cloud, local, and edge environments.

Sat, Aug 15

(3)
Context Ontology Accelerator

Context Ontology Accelerator is an open-source semantic context layer that enables AI agents to make accurate, consistent, and explainable decisions by grounding them in formal ontologies and knowledge graphs. It streamlines the process of connecting data sources, inducing business ontologies, and serving governed context to agents via the Model Context Protocol (MCP).

open
lencx
Minke

Minke is a native, local-first desktop workspace designed for DeepSeek Harness. It provides an integrated environment for agentic workflows, combining conversation management, file editing, terminal access, and plugin support into a single, cohesive application.

Ori Harness

Ori Harness is a command-line interface tool that allows developers to run existing agent CLIs on OpenRouter without changing their workflow. It provides optimized configurations, unified billing, and organization-wide guardrails for various AI agent harnesses.

Fri, Aug 14

(4)
GLM-5.3

GLM-5.3 is a frontier-grade large language model optimized for advanced coding and complex agentic workflows. It serves as the latest iteration in the GLM-5 series, delivering significant performance improvements in reasoning and code generation tasks.

Perplexity Search SDK

The Perplexity Search SDK is an agents-first Python library that enables AI agents to orchestrate custom search pipelines using 'Search as Code' primitives. By allowing agents to write Python code to control retrieval, ranking, and filtering, it reduces token usage and improves performance for complex, multi-step research tasks.

free
pgbot.dev
pgbot

pgbot is an open-source PostgreSQL intelligence tool that provides real-time database health monitoring, performance analysis, and AI-powered insights. It acts as an 'AI DBA' by analyzing database statistics to offer actionable recommendations, helping users identify and resolve performance bottlenecks without manual dashboard monitoring.

open
Qwen
Qwen3.8-27B

Qwen3.8-27B is a highly capable, open-weight multimodal model designed for advanced reasoning and agentic tasks. It features native vision and video understanding, a 256K context window, and a sophisticated thinking mechanism that can be adjusted for various performance needs.

Thu, Aug 13

(7)
paid
Arcee AI
Arcee Open Models API Beta

The Arcee Open Models API Beta expands Arcee's hosted inference service beyond its proprietary Trinity model family to include a curated selection of frontier open-weight models. This service is designed to provide developers and enterprises with greater flexibility for demanding, long-horizon agentic workloads.

CodeRabbit Security

CodeRabbit Security is an AI-powered security analysis tool that hunts for application-specific vulnerabilities across an entire codebase. It uses a four-stage workflow—Map, Hunt, Verify, and Fix—to identify complex risks like authorization bypasses and business-logic flaws, providing evidence-backed remediation directly within the development workflow.

DeepSeek Harness

DeepSeek Harness (dsh) is an open-source agent runtime framework designed for building autonomous agents. It utilizes a modular architecture where every component—including models, tools, session management, and the agent loop—is implemented as a plugin powered by the Cordis framework.

Dots3-Note Preview

Dots3-Note Preview is an open-weight, multimodal Mixture-of-Experts (MoE) model featuring 280B total parameters and 16B active parameters. Designed for long-horizon agentic tasks, it supports a 512K context window and integrates text, visual, and audio understanding to solve complex, real-world problems.

paid
Google
Gemini 3.7 Flash

Gemini 3.7 Flash is an advanced, high-performance AI model optimized for coding, agentic workflows, and complex reasoning tasks. It delivers significant improvements in debugging, issue resolution, and web development while maintaining cost-efficiency for developers and enterprises.

open
Arcee AI
NAC

NAC is an open-source agent harness designed to manage long-running, complex engineering tasks by separating temporary execution context from persistent workstream state. It utilizes a thread-and-episode architecture where a central orchestrator dispatches parallel workers to complete bounded tasks, ensuring agents maintain focus and intent over extended durations.

Toast 1

Toast 1 is a specialized search agent designed for knowledge-intensive tasks, capable of decomposing queries, gathering evidence, and curating context. It matches or outperforms frontier models like Claude Opus 5 and GPT-5.6 Sol while offering significantly higher speed and lower operational costs.

Wed, Aug 12

(4)
Delta

Delta is a multiplayer environment designed for coding with AI agents and reviewing their output. It utilizes DeltaDB to keep code and conversation threads synchronized in real-time, allowing developers and agents to collaborate with full context.

free
Docker
Docker VMM

Docker VMM is a fully rebuilt, container-optimized virtualization layer for Docker Desktop that replaces third-party solutions. It provides enhanced performance, stability, and memory management by allowing Docker to own and tune the entire virtualization stack.

paid
SpaceXAI
Grok 4.6

Grok 4.6 is a frontier AI model designed for long-running agentic tasks, complex coding, and multi-step knowledge work. It features improved reasoning capabilities and visual processing, optimized for turning broad ideas into polished applications.

open
Liquid AI
LFM2.5-VL-3B

LFM2.5-VL-3B is a 3.1-billion-parameter vision-language model designed for high-performance, on-device inference. It features significant improvements in screen understanding, grounding, and function calling, allowing it to outperform larger models while maintaining low latency on edge hardware.

Tue, Aug 11

(5)
paid
Microsoft
MAI-Code-1.1-Flash

MAI-Code-1.1-Flash is a lightweight, agentic coding model designed to provide high-quality code generation with superior efficiency. It is specifically optimized for GitHub Copilot and VS Code, offering faster token streaming and reduced operational costs for engineering teams.

open
NVIDIA
Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning 30B-A3B-NVFP4 is a high-performance, latency-optimized large language model featuring a hybrid Mamba-2 and Mixture-of-Experts (MoE) architecture. Designed for efficient agentic workflows, it utilizes 30B total parameters with 3B active parameters per token to deliver fast, accurate execution for specialized tasks.

open
NVIDIA
Nemotron-RL Agentic Terminal Pivot

Nemotron-RL Agentic Terminal Pivot v1 is an open-source reinforcement learning dataset designed to train large language models for agentic command-line interface (CLI) tasks. It provides high-quality, expert-derived trajectories that enable models to perform complex software engineering operations, tool use, and reasoning within Linux environments.

open
SolidJS
Solid

Solid 2.0 is a major update to the declarative JavaScript library for building user interfaces, introducing first-class asynchronous reactivity. By integrating async directly into the reactive graph, it eliminates the need for specialized primitives like createResource, simplifying data fetching and state management.

open
NVIDIA
Switchyard

Switchyard is a Rust-based proxy and library designed to route LLM traffic across various providers and models. It enables developers to maintain native OpenAI and Anthropic API compatibility while implementing flexible routing, benchmarking, and cost-optimization strategies.

Mon, Aug 10

(3)
Muse Glimmer

Muse Glimmer is a 30-billion-parameter open agentic model optimized for always-on local workflows on consumer hardware. It enables advanced capabilities like multi-step reasoning, reliable tool use, and multimodal understanding while running locally on devices like Macs and PCs.

open
Cohere
North Micro Vision Instruct

North Micro Vision Instruct is a 2.4B-parameter open-weight vision-language model designed for efficient multimodal tasks. It features native-resolution image processing and is optimized for prototyping, task-specific fine-tuning, and specialized visual applications.

SIE (Superlinked Inference Engine)

SIE is an open-source, self-hosted inference server designed to run a wide variety of AI models for agentic workflows, including embedding, reranking, entity extraction, and text generation. It provides a unified API that replaces fragmented model-specific servers, allowing developers to serve over 100 models from a single cluster with on-demand loading.

Sun, Aug 9

(2)
open
Microsoft
.NET Agent Skills

.NET Agent Skills is a curated collection of plugins and tools designed to provide AI coding agents with specialized domain expertise in .NET and C# development. The repository enables agents to perform complex tasks such as performance diagnostics, MSBuild optimization, Entity Framework data access, and framework migrations.

LLaDA2.2-flash

LLaDA2.2-flash is an agent-oriented Mixture-of-Experts (MoE) diffusion language model designed for long-context agentic workloads. It introduces Levenshtein Editing with DELETE and INSERT control tokens to enable efficient parallel generation, error correction, and multi-turn tool use.

Fri, Aug 7

(1)
Agent Orchestrator

Agent Orchestrator is an open-source development tool designed to manage fleets of AI coding agents in parallel. It provides a centralized dashboard to monitor agent sessions, handle git worktree isolation, and autonomously manage CI failures, code reviews, and merge conflicts.

Thu, Aug 6

(3)
open
Rei Labs
Adapt-1

Adapt-1 Preview is a non-transformer, neuro-symbolic substrate designed for test-time learning. It enables applications to learn and adapt while operating by maintaining a persistent, inspectable state without requiring task-specific pretraining or an LLM in the decision loop.

Agent Plugins

Agent Plugins is an open, vendor-neutral standard for packaging reusable components into portable plugins for AI agents. It provides a consistent format for bundling Agent Skills and Model Context Protocol (MCP) servers, allowing them to be discovered and loaded across different compatible AI agent clients.

Kitesurf

Kitesurf is a stateless, highly scalable web browser designed specifically for AI agents to perform tasks like HTML extraction and screenshot generation. It runs entirely on Cloudflare Workers, offering significantly higher efficiency in CPU and memory usage compared to traditional Chromium-based browsers.

Wed, Aug 5

(4)
Cloudflare OS

Cloudflare OS is an open-source AI productivity platform designed to function as an operating system for enterprise AI workloads. It enables employees to build, share, and run secure AI-powered applications (gadgets) while maintaining strict governance through capability-based security and Gatekeepers.

free
Meta
Muse Code

Muse Code is a terminal-based coding agent powered by the Muse Spark 1.2 model, designed to handle complex software engineering tasks across large repositories. It features persistent background agents, repository-scale execution, and built-in verification to automate planning, coding, and debugging workflows.

Prime Agent

Prime Agent is an open-source, self-improving coding harness that utilizes Recursive Language Model (RLM) and Continual Harness abstractions to manage complex, long-horizon tasks. It enables agents to programmatically manage their own state, sub-agents, and memory through a persistent IPython kernel, allowing for autonomous, iterative improvement.

open
Cursor
SDK Bridge

The Cursor SDK Bridge is an open-source protocol and local server that enables developers to drive Cursor agents from any programming language. By providing a stable sdk.v1 protobuf contract, it allows languages like Rust, Go, and Java to interact with Cursor's agent runtime without requiring direct dependency on the official TypeScript or Python SDKs.

Tue, Aug 4

(8)
open
Firecrawl
Anydoc

Anydoc is a high-performance Rust library designed to convert various office document formats, including Word, PowerPoint, Excel, and PDF, into clean, LLM-ready GitHub-Flavored Markdown. It provides consistent output across different file types and includes language bindings for Node.js and Python.

Kiro Crew

Kiro Crew is an open-source, persistent development workspace designed to transform AI coding agents into autonomous engineering teams. It enables agents to maintain context across sessions, schedule recurring tasks, and integrate with various developer tools for long-running workflows.

open
Liquid AI
LFM2.5-2.6B

LFM2.5-2.6B is a compact, agentic foundation model designed for efficient on-device deployment. It features a 128K context window and is specifically optimized for agentic workflows, including planning, tool use, and multi-step task execution.

open
DeepGrove
Maple-Preview

Maple-Preview is an open-source 20B-A1B ternary-weight reasoning large language model optimized for efficient on-device inference. It features a 24-layer, 256-expert architecture designed to deliver state-of-the-art reasoning capabilities while maintaining high performance on consumer hardware like the Mac mini M4.

paid
Pokee AI
Pokee-Isaac 28B

Pokee-Isaac 28B is a 28-billion parameter agentic model designed for long-context tasks, featuring a 10-million-token context window. It is built on a proprietary non-decoder-only architecture that allows for efficient deployment on single consumer-grade GPUs like the RTX 4090.

Qwen-MM-Plugins

Qwen-MM-Plugins is an open-source collection of multimodal skills and Model Context Protocol (MCP) servers designed to make various AI agent harnesses natively multimodal. It enables agents to perform complex tasks such as long-video analysis, 3D modeling in Blender, parametric CAD in FreeCAD, and media generation without requiring a custom runtime.

open
Vercel
Remote Agent Browser

Remote Agent Browser is a tool that enables running the agent-browser automation CLI within isolated, cloud-based Vercel Sandboxes. It allows AI agents to perform browser tasks like navigation, interaction, and data extraction in a secure, ephemeral environment.

Shieldstral

Shieldstral 1.0 is a 3B-parameter open-weights multimodal safety classifier that enables policy-adaptive content moderation. By framing moderation as a question-answering task, it allows developers to define safety policies in plain language at inference time without requiring model retraining.

Mon, Aug 3

(5)
Computer

The Computer by Cloudflare package introduces a new way to build and scale AI agents by providing them with their own dedicated 'computer' environment. Instead of relying solely on heavy, expensive containers, this runtime intelligently switches between lightweight isolates and container sandboxes based on the task requirements, such as file manipulation or native binary execution. This approach allows developers to scale agents to millions of concurrent instances while maintaining cost-efficiency and performance. By offering a durable, unified filesystem that stays in sync across different execution backends, the library simplifies the development of complex agentic systems. It provides developers with fine-grained control, auditability, and a clear paper trail of agent actions, making it a robust solution for production-grade autonomous workloads. The project is currently available as an open-source library, enabling developers to build agents that are secure, scalable, and capable of handling diverse tasks.

DeepCode

DeepCode is an open-source, multi-agent coding platform designed to automate the translation of high-level inputs—such as academic research papers, technical documentation, and natural language requirements—into production-ready software. By orchestrating specialized agents through a unified runtime, it enables complex workflows like Paper2Code reproduction, full-stack prototyping, and automated repository maintenance.

open
Octane
Octane

Octane is a high-performance JavaScript UI framework that serves as the successor to Inferno. It utilizes a compiler-first architecture to transform React-style components into direct DOM updates, eliminating the need for a virtual DOM or manual dependency array management.

Ori Eval

Ori Eval is an automated evaluation tool designed to help developers systematically identify the best AI models for their specific applications. It scans your codebase, generates test cases from your prompts, and uses an LLM-as-a-judge to grade performance across metrics like accuracy, latency, and cost.

open
OpenCore
Sandbox SDK

Sandbox SDK is an open-source TypeScript library that provides a unified interface for managing isolated code execution environments. It allows developers to spin up sandboxes across various providers, including local environments, E2B, Daytona, Vercel, Upstash, Ascii Box, and Railway, using a consistent, clean syntax.

Sun, Aug 2

(1)
paid
Alibaba
Qwen 3.8 Max

Qwen 3.8 Max is a flagship large language model developed by Alibaba, featuring a massive 2.4 trillion parameter architecture. Currently available in a preview version, it is designed to deliver frontier-level performance for complex reasoning and coding tasks.

Sat, Aug 1

(3)
OpenVuln

OpenVuln is an AI-powered tool hosted on Hugging Face that scans software repositories to identify potential security vulnerabilities. It leverages the GLM (General Language Model) family of models to assist developers in safeguarding their codebases.

paid
Fly.io
Sprites

Sprites are persistent, hardware-isolated Linux environments designed for running arbitrary code with stateful checkpoint and restore capabilities. They function as dedicated microVMs that hibernate when idle and wake automatically, making them ideal for AI agents, development environments, and secure code execution.

open
Microsoft
winapp CLI

The Windows App Development CLI (winapp CLI) is a unified command-line interface designed to streamline the development of Windows applications. It simplifies complex tasks such as managing Windows SDKs, handling MSIX packaging, generating app identity, and configuring manifests and certificates for various app frameworks.

🚀 Join the AI dev community — follow us everywhere

© 2026 MARKTECHPOST AI MEDIA INC. All rights reserved.Terms & ConditionsPrivacy Policy
Beta Mode