This paper proposes an interaction-centric taxonomy for diagnosing agent failures. Instead of labeling only the system-level outcome, it maps failures to interactions between components such as the model, harness, user, tools, memory, environment, and grader, while identifying the side responsible for repair. The taxonomy contains 41 failure modes. Its intended use is operational: model-side failures suggest post-training, harness-side failures suggest scaffolding or tool-integration changes, and environment or grader failures suggest redesigning evaluation conditions.
No heat snapshots are available in the last 24 hours.