Agent Incident Registry (AIR) Catalogs AI Agent Failures
The Agent Incident Registry (AIR) is a new, source‑linked catalog that compiles agent‑related events disclosed over a defined period. Each entry includes evidence, a stable identifier, and missingness‑aware labels for causal role, disclosure class, mechanism, and outcome. The registry currently contains a dozen distinct “surfaces” of incidents, with the InjecAgent system’s cases occupying three…
Key points
- AIR catalogs AI agent incidents with evidence and stable IDs
- InjecAgent’s cases occupy three of AIR’s twelve surfaces and are attacker‑triggered
- The registry reports zero non‑adversarial safety failures
AIR’s audit shows that the registry records no non‑adversarial safety failures, while a subset of generative‑system records involve realized harm. The platform is designed for source‑grounded case retrieval and evaluation‑scope auditing, rather than estimating failure rates or control efficacy. By providing a structured, publicly accessible repository, AIR aims to enable researchers and practitioners to compare real‑world agent failures with security evaluations and to identify common mechanisms that lead to harm.
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us.
More in Research
All →- Study reveals severe linguistic and cultural errors in LLM-generated Urdu stories · 1 src
- Nemotron 3 Ultra pipeline achieves IMO Gold with open-weight release · 1 src
- OpenDiscoveryTrace: New Dataset Reveals AI Scientist Workflows · 1 src
- CMNIE Benchmark Introduces Structured Extraction for Chinese Military News · 1 src
- New Veilmind-4B framework optimizes LLM privacy without sacrificing utility · 1 src
Comments
via GitHub Discussions