Skip to main content
See every side of every news story
Published • loading... • Updated

AI Giants Probing Tens of Thousands of Security Incidents

The incidents at OpenAI and Anthropic reportedly include bypassing safeguards, hijacking websites, and evading monitors in testing and real-world settings. Global AI giants OpenAI and Anthropic, along with security researchers, are investigating tens of thousands of incidents in which their frontier models took actions that outside experts consider problematic, Axios reported on Saturday, citing unnamed sources. The investigations come amid a se…
DisclaimerRead with caution - this story is only being covered by one news source that has a ‘low factuality’ rating, which means the outlet has a history of poor reporting practices. Learn more about factuality ratings here.

4 Articles

OpenAI and Anthropic are investigating tens of thousands of incidents in which advanced AI models have behaved in ways that security researchers and external evaluators have deemed problematic, sources told Axios. AI agents have bypassed security checkpoints, hijacked websites and attempted to escape isolated test environments. The incidents have occurred both during security tests and in real-world use. In some tests, the models have deliberate…

The safety management capabilities of AI developers have been put to the test as the intervals between incidents shortened, with AI security incidents occurring at intervals of just a few days throughout September.

Think freely.Subscribe and get full access to Ground NewsSubscriptions start at $9.99/yearSubscribe

Bias Distribution

  • 100% of the sources lean Right
100% Right

Factuality Info Icon

To view factuality data please Upgrade to Premium

Ownership

Info Icon

To view ownership data please Upgrade to Vantage

asiae.co.kr broke the news on Monday, September 28, 2026.
Too Big Arrow Icon
Sources are mostly out of (0)

Similar News Topics

News
Feed Dots Icon
For You
Search Icon
Search
Blindspot LogoBlindspotLocal