Agent Failure Atlas
Private, client-side analysis of AI-agent trace failures.
Open datasets, benchmarks, and browser tools for inspecting, replaying, evaluating, and securing AI-agent trajectories.
Private, client-side analysis of AI-agent trace failures.
Note Browser-local analysis of evidence-linked security, reliability, and control failures inside AI-agent traces.
Note Twenty deterministic synthetic sessions with expected findings, taxonomy, schemas, and evaluation metrics.
Explore and compare AI-agent trajectories in your browser
Note Browser-local inspection of checkpoints, branches, state changes, findings, and controlled alternate trajectories.
Note Ten deterministic synthetic replay scenarios with checkpoints, branches, outcomes, and comparison artifacts.
Compare memory strategies on synthetic scenarios
Note Browser-local comparison of five memory strategies across belief revision, stale procedures, duplicate evidence, privacy boundaries, temporal validity, and historical reconstruction.
Note Forty deterministic synthetic scenarios with blind strategy execution, lifecycle and lineage metrics, validity audits, and explicit limitations.