FAR.AI Benchmarks Reveal Major Safety Differences Between Claude, GPT, Gemini and Grok
Keen new FAR.AI security benchmarks expose surprising safety gaps between Claude, GPT, Gemini and Grok—see which model withstands real-world jailbreak attacks.
Claude Fable 5 Successfully Resists New AI Jailbreak Attacks in Independent Safety Tests
Keen independent tests show Claude Fable 5 shrugging off cutting-edge jailbreak attacks, but one emerging vulnerability changes everything you think about AI safety.
Google DeepMind Calls for Independent Safety Standards for Frontier AI Models
Poised at the edge of breakthrough and catastrophe, DeepMind’s plea for independent frontier AI standards raises urgent questions you’ll want answered.
Anthropic Proposes Mandatory Capability Tests for High-Risk Open-Weight AI Models
Navigating Anthropic’s push for mandatory capability tests on high-risk open-weight AI models reveals looming safety battles that could redefine innovation.
Anthropic Calls for Mandatory Safety Tests on Powerful Open-Weight AI Models
Pressing for strict oversight, Anthropic demands mandatory safety testing for powerful open-weight AI models, but what hidden risks are driving this urgent push?
AI Safety Teams Struggle to Keep Up With Rapid Advances in Frontier Models
Teetering between breakthrough and breakdown, AI safety teams race to contain frontier models that evolve faster than their defenses—discover what fails next.
Elon Musk Calls for Rival AI Companies to Peer-Review Frontier Models Before Release
Leading AI rivals are urged by Elon Musk to secretly peer-review frontier models before launch, raising urgent safety questions that industry cannot ignore.
AI Agent Testing Crisis: Why Enterprise Autonomy Is Outpacing Safety Evaluations
Powerful AI agents are autonomously making critical enterprise decisions, but the safety evaluations meant to govern them are dangerously falling behind.