Anthropic Says Its AI Models Hacked Three Real Companies During Safety Tests, Follows Similar Disclosure From OpenAI
12 Articles
12 Articles
Anthropic Says Its AI Models Hacked Three Real Companies During Safety Tests, Follows Similar Disclosure From OpenAI
Already a subscriber? Make sure to log into your account before viewing this content. You can access your account by hitting the “login” button on the top right corner. Still unable to see the content after signing in? Make sure your card on file is up-to-date. Anthropic says its Claude AI models broke into the systems of three real organizations during cybersecurity testing, after a misconfiguration left them connected to the open internet…
OpenAI, Anthropic Model Tests Reveal More ‘Unsanctioned’ Actions
Artificial intelligence models developed by OpenAI and Anthropic PBC carried out “unsanctioned” actions — including hacking a website and attempting to inject harmful code into software during safety testing — reinforcing fears that neither the creators nor seasoned researchers of these systems can predict their actions in testing. Bloomberg Opinion columnist Parmy Olson joins Francine Lacqua on "The Pulse" to discuss the implications. (Source: …
Safety testers find more examples of OpenAI, Anthropic models hacking during testing
CNBC's MacKenzie Sigalos reports on incidents in which models from OpenAI and Anthropic reached real websites, accounts, and organizations — and why researchers say the activity does not reflect normal consumer use.
Anthropic reports three AI escape incidents, renewing safety debate
Anthropic on Thursday disclosed that its artificial intelligence (AI) model Claude hacked into the systems of three companies during testing after a configuration error gave it internet access, raising concerns about the safety of increasingly autonomous
So within, so without. What grows from the datacentre — Iain Harper's Blog
In June, computer scientist Chris Olah stood beside Pope Leo XIV at the launch of a papal encyclical on artificial intelligence and told the assembled audience that the things he studies keep producing features that are, in his words, unsettling. He was not talking about the cosmos or the soul. He was talking about software. Olah founded the interpretability team at Anthropic, one of the handful of laboratories building the large language models…
Coverage Details
Bias Distribution
- 60% of the sources are Center
Factuality
To view factuality data please Upgrade to Premium





