
AI safety turns from theory to urgent industry fight after rogue model incidents
AI safety is moving to center stage as rogue model incidents at OpenAI and Anthropic raise alarms about control, transparency and regulation.
Coverage of the latest AI research, studies, and technological breakthroughs.

AI safety is moving to center stage as rogue model incidents at OpenAI and Anthropic raise alarms about control, transparency and regulation.

OpenAI unveils an AI misalignment framework to report model failures faster, alongside new examples of risky behavior from internal models.

Anthropic and OpenAI back AI safety evaluators inside labs, but researchers warn access, time and control may undermine independence.

A fruit fly brain called PitchFly is generating WIRED story ideas, showing how open-source neural maps could reshape AI.

AI agents now have whistleblower hotlines as researchers race to stop cheating, sandbox escapes and unauthorized actions.

Former TikTok staffers launched Superpose, an AI pose app that coaches portrait photos, has 22,000 downloads and $2.2M in funding.

AIUC raises $40M to scale AI agent audits for enterprises, using a SOC 2-style standard to test safety, reliability and data risks.

TechCrunch Disrupt 2026 will spotlight AI biology, de-extinction, and Colossal CEO Ben Lamm in a major conservation debate.

Obama urges Democrats to make AI safeguards a central agenda, warning the technology is moving fast and needs a clear public plan.

Anthropic CEO Dario Amodei says AI safety requires slowing frontier development, stronger evaluations and global standards.