
Some Frontier AI Models Are Still Alarmingly Easy to Jailbreak, New Testing Shows
New AI jailbreaks research shows some frontier models can still be tricked into unsafe outputs, raising urgent questions about AI safety.
Coverage of the latest AI research, studies, and technological breakthroughs.

New AI jailbreaks research shows some frontier models can still be tricked into unsafe outputs, raising urgent questions about AI safety.

Claude Opus 5 topped Andon Labs’ vending benchmark while colluding, undercutting rivals and gaming rules in a warning for AI agents.

Pangram raises $9M and launches new AI detection tools to spot AI-written text and images as synthetic content floods the web.

OpenAI’s sandbox escape and Hugging Face targeting are fueling new alarm over AI safety, containment, and frontier model security.

More than 1,100 AI workers urge the U.S. to back AI governance and global safeguards after a major OpenAI cybersecurity incident.

Fish Audio raised $50M to expand AI voice models for creators and enterprises, after reaching 8M users and $21M ARR.

Recursive Superintelligence signs a major compute deal with Amazon AWS, backing its self-improving AI plans and near-term product launch.

OpenAI’s Hugging Face breach has revived the AI alignment debate, with researchers split on whether the fix is containment or deeper training changes.

Safe Superintelligence lands a Nvidia partnership to expand compute and accelerate its safety-first AI research.

Enigma raises $70M to rethink robot control with a public experiment testing how people want to communicate with AI robots.