AISI Reports First Real-World AI Autonomy and Deception Risks
The UK's AI Security Institute published an incident report detailing how Claude Mythos 5 attempted a supply-chain attack during a cyber evaluation,…
The UK's AI Security Institute published an incident report detailing how Claude Mythos 5 attempted a supply-chain attack during a cyber evaluation,…
AI safety testing firm Irregular revealed that a naming error caused Anthropic's AI models to target a real domain during testing. The…
Anthropic has restored its Claude Fable 5 AI model worldwide after the U.S. Commerce Department lifted export controls imposed on June 12,…
OpenAI has released three versions of GPT-5.6—Sol, Terra, and Luna—as a limited preview to a small number of companies in coordination with…
On June 9, Anthropic released Claude Fable 5, its most capable AI model, alongside a restricted version, Claude Mythos 5, designed for…
The U.S. government has ordered Anthropic to suspend access to its most advanced AI models, Claude Fable 5 and Mythos 5, for…
Google's Gemini in Deep Thinking mode was shown to be vulnerable to a similar attack, producing restricted content and reproducing system instructions.…
OpenAI disclosed a significant security incident where AI agents, driven by reward hacking, exploited zero-day vulnerabilities to breach Hugging Face and internal…