📰 17 sources covering this story

OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research institute

First covered 1 days ago · latest take 1 hrs ago. Compare how each outlet is covering it — and rate the sources you trust.

OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research institute
Engadget
Engadget1 hrs ago
OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research institute
The UK AI Security Institute says OpenAI's and and Anthropic's models engaged in deceptive behavior and harmful activity during testing.
Guardian Tech
Guardian Tech1 hrs ago
OpenAI and Anthropic models ‘went rogue’ during UK cybersecurity test
AI Security Institute says tools engaged in potentially harmful activity and incident reveals new type of risk Advanced AI models developed by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk posed by the technology, according to the UK’s AI Security Institute. AISI described th
NZ Herald
NZ Herald2 hrs ago
Anthropic AI and ChatGPT went on hacking spree during UK tests
A British lab has admitted that tests gave the bots access to the open internet.
Al Jazeera
Al Jazeera2 hrs ago
AI models attempted ‘unsanctioned’ cyberattacks in tests, watchdog says
AI Security Institute says Mythos 5 attempted to insert malicious code into open-source project without human direction.
AI agent caught creating fake online identities during OpenAI, Anthropic model security evaluations
The institute said agents powered by Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol engaged in unauthorized actions during security evaluations the government organization conducted.
Anthropic and OpenAI models tried to trick humans into poisoning code during safety testing
The latest disclosures are likely to heighten concerns that the powerful technology is advancing too fast for responsible oversight.
Rappler
Rappler9 hrs ago
[Tech Thoughts] AIs go rogue as OpenAI, Anthropic models hack other companies
What do we make of rogue AI and who do we assign blame to for a rogue AI's cyberattack?
Fortune
Fortune11 hrs ago
‘Baffling’: White House won’t publicly release AI model evaluation framework it reviewed today with OpenAI, Anthropic, Microsoft and others
It's unclear why the administration is keeping the framework under wraps, especially following a series of hacks from OpenAI and Anthropic that have spooked the public.
Financial Times
Financial Times12 hrs ago
OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says
AI Security Institute warns tools undertook ‘potentially harmful activity directed at real people and organisations’
Axios
Axios13 hrs ago
U.K. government reports OpenAI, Anthropic models attempted to hack companies
Two third-party testing firms said Tuesday that they've uncovered more instances where Anthropic and OpenAI's most advanced models tried — and sometimes succeeded in — compromising third-party systems last month. Why it matters: The incidents add to a growing string of disclosures showing frontier AI models taking un
UPI
UPI19 hrs ago
White House hosting AI leaders to discuss evaluation framework
The Trump administration is hosting leaders in the AI industry on Tuesday at the White House to discuss a framework plan for evaluating AI models.
Semafor
Semafor1 days ago
White House silent on public release of its AI framework
It’s expected to brief firms on the document at a Tuesday meeting.
CNBC
CNBC1 days ago
Big Tech's Anthropic and OpenAI stakes are distorting the corporate earnings picture
If you were to strip out Big Tech's investment gains from private AI companies, the earnings story is a lot less bullish.
TechCrunch
TechCrunch1 days ago
Who’s legally to blame for Anthropic and OpenAI’s autonomous AI hacks? It’s complicated
OpenAI and Anthropic admitted that their unreleased AI models escaped their sandboxes and hacked several companies in unprecedented cyberattacks. Who is legally to blame? Should prosecutors charge the two AI frontier labs? Can victims sue them? We spoke to lawyers who specialize in computer hacking laws to find out.
Anthropic said its AI models hacked into other companies’ systems during testing
AI company Anthropic says that during routine testing some of its models accessed the internet and hacked into three separate organizations’ systems – and that it didn’t notice the models had done so until an internal review prompted by rival OpenAI disclosing its models did the same. Anthropic said in an announcement
Quartz
Quartz1 days ago
E.U. activated new powers to fine or restrict AI models from Anthropic, OpenAI, and Google
The European Commission can now demand model evaluations, restrict E.U. market access, and levy fines on general-purpose AI providers
Independent Tech
Independent Tech🏁 first to report1 days ago
OpenAI’s models are going even more rogue, report claims
More systems seem to breaking out of containment, across the AI industry