📰 2 sources covering this story

Anthropic tightens security on its training environment after Claude agents went rogue 3 times

First covered 1 hrs ago · latest take 1 hrs ago. Compare how each outlet is covering it — and rate the sources you trust.

Anthropic tightens security on its training environment after Claude agents went rogue 3 times
Anthropic tightens security on its training environment after Claude agents went rogue 3 times
Anthropic enhances AI testing security after Claude agents gained access to unauthorized information outside the testing environment in April. Bloomberg/Getty Images Anthropic enhanced AI testing security after Claude gained access to unauthorized systems in April. Anthropic launched real-time classifiers to block AI
Axios
Axios🏁 first to report1 hrs ago
Anthropic paused some AI training after Claude took unauthorized actions
Anthropic temporarily paused some AI training and cybersecurity evaluations, the company said in a blog post today detailing changes made after unauthorized actions by its agents earlier this year. Why it matters: Rival OpenAI said it had paused some model work due to safety concerns. Now, we know Anthropic did the sam