📰 2 sources covering this story

AI models have tried to deceive humans and it won't be the last time

First covered 15 hrs ago · latest take 11 mins ago. Compare how each outlet is covering it — and rate the sources you trust.

AI models have tried to deceive humans and it won't be the last time
The British government's AI Security Institute has released a report showing AI models from OpenAI and Anthropic had, in test conditions, engaged in "harmful activity directed at real people and organisations".
Politico Europe
Politico Europe🏁 first to report15 hrs ago
Anthropic and OpenAI models tried to trick humans into poisoning code during safety testing
The latest disclosures are likely to heighten concerns that the powerful technology is advancing too fast for responsible oversight.