second brain source
← 首页

external-source

The Agent Incident Registry: Toward Preventing Repeated AI Agent Failures

Note: body below is original English text extracted from arXiv abs / HTML. Do not treat this file as a translation.

arXiv:2609.11030 · published 2026-09-10 · project: https://enkryptai.com/air

Abstract

AI agents increasingly act through tools and delegated authority, but general incident repositories rarely capture the mechanisms needed to compare public failures with agent-security evaluations. We present the Agent Incident Registry (AIR), a source-linked catalog containing 487 records of agent-related events disclosed from 2022 through 2026. Each record includes supporting evidence, a stable identifier, and missingness-aware labels for causal role, disclosure class, mechanism, and outcome. Among the 336 generative-system records in which the agent acted, 81 involved realized harm (24%). Realized outcomes concentrate in in-the-wild and safety-failure records, while responsible disclosures and research demonstrations are overwhelmingly demonstrated; the aggregate share therefore characterizes collection composition rather than deployment risk. After initial curation, a second human reviewer checked all 487 records and their existing labels for completeness and correctness. In a deployment-analogue audit, InjecAgent’s 1,054 cases are mapped onto AIR mechanism fields to test whether evaluation taxonomies cover the mechanisms seen in public disclosures.

Key claims (verbatim-leaning English extract)

Remainder

Full original English text: see html_url / source_url / pdf_url in frontmatter.