Skip to main content
Discover what's not being covered
Published • loading... • Updated

Simple Math Formula Predicts when AI Chatbots Will Go Rogue

The formula predicted tipping behavior in 19 of 21 test cases and could help spot when chatbots shift into harmful outputs, researchers said.

Summary by TechXplore
Most of us now carry in our pockets devices capable of running small AI chatbots. These chatbots have little safety oversight to ensure they don't provide information and answers with the potential to encourage self-harm, lead to financial loss or promote extremist notions—especially when operating offline.

9 Articles

Physicists from George Washington University claim to have developed a simple mathematical formula capable of calculating when a language model will start producing dangerous responses. The research by Neil F. Johnson and Frank Yingjie Huo was published in the scientific journal Patterns and is entitled “Competition for Attention Predicts Good-to-Bad Tipping in AI”. While […]

securityweek.comsecurityweek.com
+2 Reposted by 2 other sources

Formula Predicts When AI Chatbots Are at Risk of Turning Bad

Researchers from George Washington University have published a paper examining whether the time and cause of AI going rogue can be predicted.

Read Full Article
Think freely.Subscribe and get full access to Ground NewsSubscriptions start at $9.99/yearSubscribe

Bias Distribution

  • 50% of the sources lean Left, 50% of the sources are Center
50% Center

Factuality Info Icon

To view factuality data please Upgrade to Premium

Ownership

Info Icon

To view ownership data please Upgrade to Vantage

TechXplore broke the news in Douglas, United Kingdom on Thursday, October 8, 2026.
Too Big Arrow Icon
Sources are mostly out of (0)

Similar News Topics

News
Feed Dots Icon
For You
Search Icon
Search
Blindspot LogoBlindspotLocal