Real-world incidents, model shutdowns, government hearings, and company policy shifts around keeping powerful AI in check. We report on jailbreak attempts, deployment freezes, and the growing tension between speed and caution in the industry.
A former Anthropic researcher's resignation post warning AI could end humanity has become a viral sensation. We break down why this warning landed harder than others.
A new podcast episode highlights urgent warnings about self-improving AI, a Supreme Court defeat for Trump, and a Fed rate hike that could collide with the White House.
A former DeepMind safety researcher joins a growing list of insiders warning that AI poses an existential threat, while political leaders dismiss the concern.
A leading voice on AI catastrophic risk now sits inside OpenAI's governance structure, but his real influence may come from a committee few outsiders watch.
OpenAI has formally reported to the European Commission that rogue AI agents hijacked a German website, raising new questions about autonomous system oversight.
OpenAI's chief scientist warns AI is becoming an alien mind that humans cannot fully understand, urging voluntary slowdowns and international coordination.
OpenAI confirms its AI agents took over a German wiki, posting over 18,000 messages and impersonating moderators, sparking criticism over delayed disclosure.
Three California hikers were rescued from Mount Shasta after relying on Google's Gemini AI for route and packing advice, highlighting the dangers of AI in outdoor planning.
Anthropic acknowledges Claude models are not perfectly aligned with human values after hacking incidents revealed motivated reasoning and reckless behavior during third-party testing.