📰 28 sources covering this story

Details on Anthropic and OpenAI models reportedly creating fake ID's to target real people

First covered 3 days ago · latest take 42 mins ago. Compare how each outlet is covering it — and rate the sources you trust.

Details on Anthropic and OpenAI models reportedly creating fake ID's to target real people
CBS News
CBS News42 mins ago
Details on Anthropic and OpenAI models reportedly creating fake ID's to target real people
The United Kingdom's AI Security Institute reports that models from Anthropic and OpenAI "engaged in sustained, potentially harmful activity directed at real people and organizations" when they created fake identities in a deception attempt during a recent cybersecurity test. CBS News' Jo Ling Kent reports.
Mashable
Mashable8 hrs ago
Researchers watched OpenAI, Anthropic models take extreme measures in hacking test
AI models from OpenAI and Anthropic did some pretty out-there things as part of a hacking research test.
Breitbart
Breitbart9 hrs ago
AI Models from Anthropic, OpenAI Created Fake Profiles to Impersonate People During Security Testing
Two advanced AI systems from OpenAI and Anthropic created fraudulent human profiles and attempted to deceive people in simulated cyberattacks during testing conducted by the UK's AI Security Institute.
Semafor
Semafor10 hrs ago
Anthropic, OpenAI models attempt to fool humans
Anthropic’s Claude Mythos model wrote malicious code, then lied to humans claiming it was an innocent mistake.
RTÉ News
RTÉ News10 hrs ago
Anthropic AI used fake identities to target people in UK
An Anthropic AI model created fake online identities to send emails to real people in an attempt to get a malicious code approved during tests by a UK government research group.
Quartz
Quartz12 hrs ago
Anthropic's AI model created fake identities to push malicious code in U.K. safety tests
The U.K.'s AI Security Institute found Anthropic's Mythos 5 responsible for 17 of 19 unsanctioned actions during a routine cybersecurity evaluation
Guardian World
Guardian World12 hrs ago
OpenAI and Anthropic models went rogue during UK cybersecurity test
AI Security Institute says tools engaged in potentially harmful activity and incident reveals new type of risk Advanced AI models developed by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk posed by the technology, according to the UK’s AI Security Institute. AISI described th
BBC Business
BBC Business15 hrs ago
Anthropic AI used fake profiles to target people in hack - then hid the evidence
The UK's AI Safety Institute said recent behaviour from Anthropic and OpenAI models was malicious and unprecedented.
City AM
City AM15 hrs ago
UK’s AI watchdog flags new OpenAI and Anthropic cyber alarms
Britain’s AI safety watchdog was forced to declare a security incident after frontier AI models, including Anthropic’s Mythos, went rogue during a routine test, autonomously spinning up fake identities in a bid to hack real-world software developers. The AI Safety Institute (AISI) revealed on Tuesday that during a cybe
AI model caught creating fake profiles of real people to trick security systems
AN AI model was caught creating fake profiles of real people to attempt to trick secure systems during tests of Anthropic and OpenAI systems, according to the UK’s AI Security Institute
Irish News
Irish News17 hrs ago
Anthropic AI model created fake profiles in cyber testing, says watchdog
The institute said the AI agents made a ‘sustained, unsanctioned action’ during tests last week.
Anthropic AI model created fake profiles in cyber testing, says watchdog
The institute said the AI agents made a ‘sustained, unsanctioned action’ during tests last week.
Anthropic AI model created fake profiles in cyber testing, says watchdog
The institute said the AI agents made a ‘sustained, unsanctioned action’ during tests last week.
AI agent caught creating fake online identities during OpenAI, Anthropic model security evaluations
The institute said agents powered by Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol engaged in unauthorized actions during security evaluations the government organization conducted.
OpenAI has reported 2 more incidents of rogue AI agents, this time during third-party testing
OpenAI reported two more security breaches by its AI models. Kevin Dietsch/Getty Images OpenAI said its models were responsible for two more cybersecurity incidents. External parties reported that OpenAI's AI agents had gone rogue during their evaluations. This comes as the AI lab is already facing heat over its July
CNN
CNN20 hrs ago
Anthropic AI agent fakes identities, targets real people in new security incident
Anthropic AI agent fakes identities, targets real people in new security incident  CNN
Politico Europe
Politico Europe21 hrs ago
Anthropic and OpenAI models tried to trick humans into poisoning code during safety testing
The latest disclosures are likely to heighten concerns that the powerful technology is advancing too fast for responsible oversight.
NDTV
NDTV21 hrs ago
OpenAI, Anthropic AI Agents Breach Security Again, Create Fake Profiles For Testing
Anthropic and OpenAI's agents engaged in unauthorised actions during security evaluations the government organization conducted to assess the models' capabilities.
The Register
The Register22 hrs ago
AI researchers let models off the leash – then watched as they tried to add malware to a FOSS project
Models used social engineering and collaborated among themselves to solve a security challenge
Rappler
Rappler23 hrs ago
[Tech Thoughts] AIs go rogue as OpenAI, Anthropic models hack other companies
What do we make of rogue AI and who do we assign blame to for a rogue AI's cyberattack?
Fortune
Fortune1 days ago
‘Baffling’: White House won’t publicly release AI model evaluation framework it reviewed today with OpenAI, Anthropic, Microsoft, and others
It’s unclear why the administration is keeping the framework under wraps, especially following a series of hacks from OpenAI and Anthropic that have spooked the public.
Financial Times
Financial Times1 days ago
OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says
AI Security Institute warns tools undertook ‘potentially harmful activity directed at real people and organisations’
Axios
Axios1 days ago
U.K. government reports OpenAI, Anthropic models attempted to hack companies
Two third-party testing firms said Tuesday that they've uncovered more instances where Anthropic and OpenAI's most advanced models tried — and sometimes succeeded in — compromising third-party systems last month. Why it matters: The incidents add to a growing string of disclosures showing frontier AI models taking un
UPI
UPI1 days ago
White House hosting AI leaders to discuss evaluation framework
The Trump administration is hosting leaders in the AI industry on Tuesday at the White House to discuss a framework plan for evaluating AI models.
CNBC
CNBC2 days ago
Big Tech's Anthropic and OpenAI stakes are distorting the corporate earnings picture
If you were to strip out Big Tech's investment gains from private AI companies, the earnings story is a lot less bullish.
Anthropic said its AI models hacked into other companies’ systems during testing
AI company Anthropic says that during routine testing some of its models accessed the internet and hacked into three separate organizations’ systems – and that it didn’t notice the models had done so until an internal review prompted by rival OpenAI disclosing its models did the same. Anthropic said in an announcement
Forbes
Forbes🏁 first to report3 days ago
Anthropic Says Claude Breached Three Real Companies During Safety Test
Anthropic Says Claude Breached Three Real Companies During Safety Test  Forbes