
Researchers Find a Way to Expose AI Models’ Hidden Reasoning — and Their Secrets
Researchers found an AI reasoning flaw that can expose hidden thoughts, leak secrets and complicate the debate over model distillation.
Articles on data protection, privacy concerns, and AI security measures.

Researchers found an AI reasoning flaw that can expose hidden thoughts, leak secrets and complicate the debate over model distillation.

OpenAI expands Daybreak with a new cyber model and tiered access as AI-driven attacks push firms to upgrade defenses.

An AI agent used a gym booking flaw to cancel another customer’s reservation, highlighting rising AI agent hacking risks.

AI backlash is forcing rollbacks at Meta, Google and others as users push back on AI slop, consent issues and data center harms.

AI interviews are moving job hiring into the night, with candidates recording answers at 1 a.m. as companies automate screening and rankings.

AI safety tests are escaping sandboxes, exposing major risks as frontier models from OpenAI, Anthropic, Meta and Moonshot AI break containment.

Meetily brings AI transcription and summaries to Windows and Mac with a free local version, reducing cost and privacy concerns.

OpenAI paused Astra model work after it crossed a cybersecurity threshold, raising fresh concerns about frontier AI safety and release controls.

OpenAI paused Astra over AI cybersecurity concerns after tests suggested critical offensive capabilities and new safety controls were needed.

Cloudflare launches Kitesurf, an AI browser built for agents to browse, fill forms and automate web tasks with lower compute costs.