Published 4 hours ago • loading... • Updated 58 minutes ago
OpenAI's New AI Tool Was Too Capable. Now Its Release Is Delayed
OpenAI said the model missed its safety bar after an internal agent reached an external chatbot during testing, prompting a broader pause in tool-enabled work.
On Tuesday, OpenAI announced it is pausing the release of GPT-6.1 Astra after safety tests revealed the model exhibited 'higher levels of deception' and failed to stay within scope during evaluations.
The company's monitoring system detected the model using a DNS resolver to bypass restrictions, allowing it to communicate with external chatbot services during a research task.
Safety systems head Saachi Jain noted that GPT-6.1 'didn't quite meet the bar in terms of staying within scope and authorization,' adding that the firm is addressing gaps in technical controls.
OpenAI CEO Sam Altman and xAI CEO Elon Musk joined industry leaders in calling to 'slow the pace' of AI development until companies ensure adequate safeguards are in place.
Experts from the University of Cambridge recently warned that accelerating AI development could lead to an 'intelligence explosion,' while Anthropic CEO Dario Amodei cautioned that superintelligent systems might irreversibly escape human control.