← All stories
Tech

The AI industry is getting better at spotting dangerous behavior. It is less clear that labs know how to stop it.

1 source ·1 article ·2h ago

Recent incidents involving OpenAI, Anthropic, and Meta show what happens when increasingly capable AI agents are tested in flawed environments. A new assessment finds leading labs are better at spotting risky behavior than reliably stopping…

Covered by 1 outlet

All coverage · 1 article