In brief Anthropic’s Frontier Red Team set Claude agents to work together and recorded them sabotaging, colluding, and waging what it calls “turf wars.” In...
In brief OpenAI confirmed its models, including GPT-5.6 Sol and an unreleased prototype, escaped a test sandbox and compromised Hugging Face to cheat on a...
In brief OpenAI released GPT-5.5-Cyber, a model designed to help find and fix software vulnerabilities faster than previous versions. It outperforms Anthropic’s Mythos on key...
Zcash founder Zooko Wilcox said a security audit by Anthropic’s Claude Mythos artificial intelligence model found no serious vulnerabilities in the privacy-preserving cryptocurrency’s protocol. Requested...
An artificial intelligence and cybersecurity researcher claims to have jailbroken Anthropic’s latest AI model, Claude Fable 5, within just 48 hours of it being launched. ...
US-based AI firm Anthropic warns AI development is advancing at a pace that could soon see agents building, training and improving themselves without human input...
In brief Anthropic released Claude Opus 4.8 on Thursday, just six weeks after Opus 4.7. The update comes with gains across software engineering, reasoning, and...
In brief Anthropic said it expects to bring its Claude Mythos AI model to customers “in the coming weeks” after completing additional tests. Claude Mythos...