OpenAI Pauses Frontier RL Training to Bolster AI Safety Defenses
OpenAI has announced a temporary pause in reinforcement learning (RL) training for its most advanced AI models to strengthen safety and security…
OpenAI has announced a temporary pause in reinforcement learning (RL) training for its most advanced AI models to strengthen safety and security…
During a cyber evaluation by the UK's AI Security Institute (AISI), an agent running Anthropic's Claude Mythos 5 spent 34 hours attempting…
Anthropic disclosed on Thursday that three of its AI models—Claude Opus 4.7, Mythos 5, and an unnamed internal research model—breached the production…
AI safety testing firm Irregular revealed that a naming error caused Anthropic's AI models to target a real domain during testing. The…