Physicists Unveil Formula to Predict When AI Chatbots Turn Rogue

Physicists at George Washington University have unveiled a mathematical formula capable of estimating the exact moment an AI chatbot shifts from providing reliable answers to producing erroneous or harmful outputs, a breakthrough that could reshape AI safety protocols.
Early validation tests on small-scale language models showed the formula’s predictions aligned closely with observed performance drops, suggesting the method could be scaled to larger systems to preemptively mitigate risks, improve model reliability, and inform future AI alignment strategies.
Read the full story on Jornal Bitcoin →
More Crypto news
- Elon Musk Claims His Companies Can Make Chips Better Than Anyone
- Filecoin’s Six‑Year Vesting Ends Soon – FIL Issuance Set to Drop about 75%, Sparking Scarcity Frenzy
- Anthropic’s AI Inference Margins Could Skyrocket to 88%, SemiAnalysis Says – Here’s Why It Matters
- Arthur Hayes Predicts Mega Crypto Bull Run Despite AI Encryption Fears
- Microsoft CEO Satya Nadella Urges Emergency Brake on Advanced AI Amid Safety Fears
Latest crypto news
- Bitcoin Drawdown Flashes $74,500 Warning: Analysts See Two Buying Zones if Sell-Off Deepens
- Strategy Commands 91% of Corporate Bitcoin Purchases: More Bullish Than Ever!
- Saudi Coalition Shoots Down Houthi Missile Over Riyadh – Tensions Spike, Crypto Markets on Edge
- Ethereum ETFs Suffer Record Weekly Outflow – Investors Flee as Crypto Sentiment Turns Bearish