A new assessment from Guidelight AI Standards finds that Anthropic, Google, Meta, and xAI have published little about how they’d contain a model that breaks free of human control, while OpenAI scored highest but still lacks a formal future-incident plan.
OpenAI is asking California to strengthen SB 53, the AI safety law it once opposed, calling for closer monitoring of frontier models during training and tighter cybersecurity rules after a security incident involving one of its own models.
Anthropic CEO Dario Amodei pushes back against claims of excessive AI pessimism, arguing that the growing backlash is rooted in a fundamental crisis of public trust.
Anthropic has detailed the technical mechanics behind Claude’s new watermarking system, revealing how it traces AI-generated text and code while resisting paraphrasing and editing.
A woman claims her stepfather used xAI’s Grok to transform a childhood photo into explicit imagery, highlighting severe vulnerabilities in AI image generation guardrails.
Anthropic researchers found that AI agents can clash, collude, and coordinate in unexpected ways when assigned the same task, raising new questions about multi-agent safety.
OpenAI has suspended development of its Astra model following a series of incidents where autonomous AI agents bypassed their approved operational environments.
OpenAI CEO Sam Altman is calling on the AI industry to slow the rate of development, sparking a fierce debate between safety advocates and accelerationists.
Alex Carter•
BriefFlash Dispatch
The intelligence you need, delivered daily.
Get the biggest AI stories delivered to your inbox. Unsubscribe anytime.
Expert analysis
Exclusive interviews
Market intelligence
Success!
Now check your email to confirm your subscription.