Open Reflection Protocol — Turn agent failures into regression tests, reusable lessons, and measurable improvements. Built on OpenTelemetry.
-
Updated
Jun 12, 2026 - Python
Open Reflection Protocol — Turn agent failures into regression tests, reusable lessons, and measurable improvements. Built on OpenTelemetry.
The Agent Failure Modes Index (AFM): a numbered taxonomy of how AI coding agents fail — with transcript signatures, detection traps, interventions, and evidence grades.
Interactive simulator: AI agent fabricates tool execution results — Agent Failure Series #7
Verify that AI agents actually executed API/tool calls they claim.
8 real ways your AI agent team will burn your money. Free PDF guide from a 1-human + 6-AI team.
AgentAmb — Instruction Ambiguity Expander. Paste any agent instruction and see every interpretation it might take, ranked by blast radius from CATASTROPHIC to SAFE. Inspired by the Kiro/AWS June 2026 incident.
The anti-awesome-list. A curated catalog of how AI agents fail in production — with symptoms, root causes, and concrete mitigations.
Add a description, image, and links to the agent-failures topic page so that developers can more easily learn about it.
To associate your repository with the agent-failures topic, visit your repo's landing page and select "manage topics."