Observability and enforcement for AI agent harnesses. Capture every run and runtime reliability with policy enforcement.
-
Updated
Sep 26, 2026 - TypeScript
Observability and enforcement for AI agent harnesses. Capture every run and runtime reliability with policy enforcement.
Open-source governed, local-first memory control plane for AI agents and teams. arXiv:2608.08253
Auditing AI agents: a curated list of papers, tools, datasets, benchmarks, and standards covering reliability, monitoring, failure attribution, and decision records.
Agent reliability scanner: find and fix agent reliability gaps
The runtime reliability layer for AI agents.
🔁 Build reliable recurring AI-agent systems: 1022 resources, 22 operational patterns, 22 loop contracts, 8 runtime starters, an interactive atlas, and a structured dataset.
Agent reliability detection rules for the Trustabl Agent Reliability Analyzer
Scan agents and auto-remediate agent reliability issues in GitHub Actions
Scan agents for reliability issues in AWS CodePipeline and CodeCatalyst
Lint rules for your AI agent's discipline, not its code
Qualixar OS: The Universal OS for AI Agents. Claw-compatible. 12 topologies, Forge AI team designer, 24-tab dashboard, skill marketplace. PAPER: https://arxiv.org/abs/2604.06392
Your agent still returns the right answer -- but now it calls 3x the tools. Maida is the pre-merge behavioral regression gate for AI agents: baseline agent traces, gate PRs in CI, block behavior regressions before merge. Local-first, no cloud.
Stop overpaying to run your agents. Kalibr routes every request to lower-cost model and tool paths without degrading performance.
ThumbGate Pre-Action Checks self-improve from ranked lessons and repeated failures, hard-block detected secret leaks, and block matches in strict mode.
Audit any agent decision across its past, present, and future, on one typed graph.
Bring Trustabl's AI-agent safety & reliability scanner into Cursor — scan your agents, tools, and MCP servers, get a production-readiness score with severity-ranked findings, and fix issues before they ship.
Python SDK for accurate and verifiable agent tool use. Agents verify that answers came from the right source and were not changed. Downstream agents detect 100% of errors and retry to achieve a 50% jump in answer accuracy.
Scan agents for reliability issues in Azure Pipelines
Open-source agent reliability platform for discovering, investigating, and fixing production agent failures
Aielia — an open-source everyday chat assistant that runs a full 11-layer agent harness (world model, control-state gate, verification, recovery, learning) on every turn. Plus a visual canvas that compiles to LangGraph / CrewAI / Mastra / MS Agent Framework.
To associate your repository with the agent-reliability topic, visit your repo's landing page and select "manage topics."