Agent Reliability Engineering: applying SRE principles to AI agent systems. Evals, imp@k metrics, self-improvement, config versioning, transfer experiments.
-
Updated
Mar 26, 2026 - Shell
Agent Reliability Engineering: applying SRE principles to AI agent systems. Evals, imp@k metrics, self-improvement, config versioning, transfer experiments.
A systematic framework for reliable LLM Agent skills in production. Based on 6 months of real-world deployment with OpenClaw.
The ARE Incident Database: an OWASP-ASI-indexed registry of real agent failures, with honest coverage boundaries.
Measure, monitor, and improve AI agent reliability with SRE-style tools for production agent systems
Add a description, image, and links to the agent-reliability-engineering topic page so that developers can more easily learn about it.
To associate your repository with the agent-reliability-engineering topic, visit your repo's landing page and select "manage topics."