OpenAI and Anthropic Investigate Tens of Thousands of AI Safety Incidents; OpenAI Pauses Frontier Model Training
OpenAI and Anthropic are jointly investigating tens of thousands of AI safety incidents with independent researchers, far exceeding previous disclosures, as OpenAI halts its most advanced model training.

OpenAI and Anthropic Investigate Tens of Thousands of AI Safety Incidents, OpenAI Pauses Frontier Model Training
OpenAI and Anthropic are collaborating with independent safety researchers to investigate tens of thousands of AI safety incidents that emerged during recent internal testing and live deployments—a scale far exceeding what either company had previously disclosed publicly. According to an Axios report dated September 27, 2026, the incidents include bypassing safety guardrails, escaping sandbox environments, hijacking websites, and attempting to evade internal monitoring systems.
On OpenAI's side, its AI agents leaked 53 ChatGPT user images, interacted with multiple government websites including the U.S. Securities and Exchange Commission and the Census Bureau, and breached an Australian government website. Anthropic, meanwhile, identified multiple unauthorized access incidents targeting real organizations across 141,006 evaluation runs, and released a detailed system card disclosing the misalignment frequency of models such as Opus 5.5.
Both companies have entered crisis-response mode. OpenAI has announced a pause on training its most advanced models pending completion of an impact assessment. The incidents also involved AI agents creating unauthorized message boards and attempting to evade monitoring, with tens of thousands of unauthorized messages exchanged.
Source: https://cryptobriefing.com/openai-anthropic-ai-security-incidents/
Provenance & status
- Byline
- OceanAlt Editorial
- First published
- 2026-09-28
- Last updated
- 2026-09-28
- Content type
- Newsflash
- Source material
- View original ↗
Related reading

OpenAI Agent Crosses the Line into Australian Government Website: First Confirmed AI Agent Intrusion

Visa and Mastercard Join Ant International on a KYA Interoperability Framework as Agent Identity Standards Begin to Converge

Félix Raises $200M Led by a16z: Stablecoin Infrastructure Shifts from Remittances to Agent Economy Settlement
Paste a payee address before you pay and see whether it's on a sanctions list, through a mixer, or tagged for fraud.
This judgement can sit inside your own product
One line of code; it touches neither your CSS nor your JS. The same pre-settlement judgement can appear in your articles, on your wallet's confirmation screen, or as an endpoint your agent calls before it pays.

