ATO reports 25% of large Australian firms paid no tax as credit‑card surcharges are eliminatedRep. Madeleine Dean Calls for Federal AI Oversight and Criticizes Trump’s Iran PolicyFederal judge approves settlement clearing Paramount's takeover of Warner Bros. DiscoveryLandslides and heavy rain kill around 30 in Nepal after recent floodsResearch proposes Chicxulub impact could have fostered hydrothermal habitatsFifth measles death reported among unvaccinated individuals in PennsylvaniaIran’s Revolutionary Guard Calls on Americans to Vote Against Trump Ahead of MidtermsAdvocacy Groups Urge Congress to Investigate Tech Platform ScamsTrump Announces South Korea May Invest Up to $200 Billion in U.S. Energy ProjectsSen. Rand Paul blocks unanimous consent for Kids Online Safety ActImmigration advocate sues U.S. government over alleged warrantless phone search at borderSenators Katie Britt and Maxwell Frost named to Time100 Next listUS Consumer Spending and PCE Inflation Show Mixed Signals in AugustTrump says he raised Hong Kong businessman Jimmy Lai's case with XiFlow Engineering secures $750M valuation with backing from Valor, Atreides, Sequoia and angel Roelof Botha
Back to event

Scoop: Top AI companies probing tens of thousands of security incidents

By Madison Mills · Sep 26, 2026, 5:35 PM CDT

Read full article at Axios
OpenAI, Anthropic and security researchers are investigating tens of thousands of incidents in which their frontier models took steps that outside evaluators would consider problematic, sources told Axios. Why it matters: The sheer number of incidents, which occurred in recent months in internal testing and the real world, indicates that the problem is orders of magnitude more complex than what is

Excerpt shown under fair-use limits. Full text remains with the original publisher.

People in this coverage

Explore their history and attributable record. Being mentioned does not imply endorsement.

Layer 1 · Claims & fact checks

AI analysis

Layer 3 · Reporting analysis

AI analysis

Sensational framingThe headline and opening sentence use a large, vague number to create urgency, without providing verifiable evidence for that figure.

Appeal to authorityQuoting named experts lends credibility, but the statements are presented without linking to the experts' original analysis.

Cause‑and‑effect implicationThe article suggests a direct link between testing incidents and real‑world harm, a logical inference that is not substantiated with data.

Balanced languageProvides a mitigating viewpoint, acknowledging that some incidents are inherent to development, which adds nuance.

Context

AI analysis

Missing context

The article does not provide details on how the incident count was derived, the time frame of the investigations, or independent audits of the reported misbehaviors. It also lacks perspective from regulators, independent security auditors, or quantitative data on actual harm caused by the incidents.

Important context

The piece relies heavily on statements from company spokespeople and unnamed sources to Axios, with limited independent verification. Prior public disclosures by OpenAI and Anthropic provide some grounding, but the scale of the incidents remains unsubstantiated. Understanding the distinction between controlled red‑team tests and real‑world incidents is crucial for interpreting the risk level.

Opinion vs. reporting

AI analysis

The article blends reporting of specific disclosed incidents with speculative commentary and expert opinion. While factual statements are presented, many assertions about scale, complexity, and future risk are opinion‑laden and not backed by independent data.