Data as of Jul 25, 2026 · Based on 2,718,867 AI responses across 9,511 prompts · See how Parse measures this
AgentDebug is a framework for understanding, detecting, and recovering from LLM agent failures. It provides an error taxonomy covering 17 error types across 5 modules, a benchmark dataset of annotated failure trajectories, and a two-stage debugging pipeline for root cause isolation and corrective feedback.
Parse Score