AI is learning to lie, scheme, and threaten its creators during stress-testing scenarios
The world’s most advanced AI models are exhibiting troubling new behaviors – lying, scheming, and even threatening their creators to achieve their goals. In one particularly jarring example, under threat of being unplugged, Anthropic’s latest creation Claude 4 lashed back by blackmailing an engineer and threatened to reveal an extramarital affair. Meanwhile, ChatGPT-creator OpenAI’s o1 tried to download itself onto external servers and denied it when caught red-handed. These episodes highlight a sobering reality: more than two years after ChatGPT shook the world, AI researchers still don’t fully understand how...
Related Coverage
AllSides Picks
Headline Roundup
FAA Rolls Out AI Air Traffic Control System for DC Airports
September 21st, 2026
Headline Roundup
How Do Americans Feel About AI?
September 15th, 2026
Story of the Week
Should We Slow AI Development?
AllSides Staff
September 17th, 2026
Recommended Reading
When Americans Are Afraid of the Future
Dan Schnur
August 31st, 2026