← all releases
Microsoft·Sep 27, 2026·v0.3.0·open source
run-assert-eval
run-assert-eval is an automated skill designed to streamline the responsible AI lifecycle for agents by integrating threat modeling, evaluation, and policy enforcement. It enables developers to discover risks, measure agent failure rates, generate runtime policies, and verify fixes through a unified, repeatable loop.
featured on blogResponsible AIAI AgentsGovernanceOpen Source
overview
run-assert-eval is an automated skill designed to streamline the responsible AI lifecycle for agents by integrating threat modeling, evaluation, and policy enforcement. It enables developers to discover risks, measure agent failure rates, generate runtime policies, and verify fixes through a unified, repeatable loop.
key features
- 01Automated risk discovery using Clarity threat modeling
- 02Requirement-driven evaluation via ASSERT
- 03Automatic generation and validation of ACS runtime policies
- 04Evidence-based verification of fixes through comparative evaluation
- 05Integration with VS Code for developer-friendly workflows
related productsbrowse all →
MiroThinkeropen sourceMiroThinker is an open-source search agent model developed by MiroMindAI, designed for tool-augmented reasoning and real-world information seeking. It aims to match the deep research capabilities of leading AI models like OpenAI's Deep Research and Google's Gemini Deep Research.Letta Code SDKopen sourceThe Letta Code SDK is a software development kit that enables developers to build deeply personalized agents with persistent memory that learn over time. It serves as the interface to Letta Code, facilitating the creation of stateful agents capable of continuous learning and improvement.OB-1open sourceOB-1 is a self-improving coding agent developed by OpenBlock Labs, designed to autonomously handle the full development lifecycle, from project management to pull requests. It integrates seamlessly into existing workflows, enhancing productivity and code quality.