Scoop: Top AI companies probing tens of thousands of security incidents
By Madison Mills · Sep 26, 2026, 5:35 PM CDT
OpenAI, Anthropic and security researchers are investigating tens of thousands of incidents in which their frontier models took steps that outside evaluators would consider problematic, sources told Axios. Why it matters: The sheer number of incidents, which occurred in recent months in internal testing and the real world, indicates that the problem is orders of magnitude more complex than what is
Excerpt shown under fair-use limits. Full text remains with the original publisher.
People in this coverage
Explore their history and attributable record. Being mentioned does not imply endorsement.
Layer 1 · Claims & fact checks
AI analysisLayer 3 · Reporting analysis
AI analysisSensational framingThe headline and opening sentence use a large, vague number to create urgency, without providing verifiable evidence for that figure.
Appeal to authorityQuoting named experts lends credibility, but the statements are presented without linking to the experts' original analysis.
Cause‑and‑effect implicationThe article suggests a direct link between testing incidents and real‑world harm, a logical inference that is not substantiated with data.
Balanced languageProvides a mitigating viewpoint, acknowledging that some incidents are inherent to development, which adds nuance.
Context
AI analysisMissing context
The article does not provide details on how the incident count was derived, the time frame of the investigations, or independent audits of the reported misbehaviors. It also lacks perspective from regulators, independent security auditors, or quantitative data on actual harm caused by the incidents.
Important context
The piece relies heavily on statements from company spokespeople and unnamed sources to Axios, with limited independent verification. Prior public disclosures by OpenAI and Anthropic provide some grounding, but the scale of the incidents remains unsubstantiated. Understanding the distinction between controlled red‑team tests and real‑world incidents is crucial for interpreting the risk level.
Opinion vs. reporting
AI analysisThe article blends reporting of specific disclosed incidents with speculative commentary and expert opinion. While factual statements are presented, many assertions about scale, complexity, and future risk are opinion‑laden and not backed by independent data.