📰 16 sources covering this story

OpenAI, Anthropic AI agents implicated in new security breaches

First covered 2 days ago · latest take 1 hrs ago. Compare how each outlet is covering it — and rate the sources you trust.

OpenAI, Anthropic AI agents implicated in new security breaches
Breitbart
Breitbart3 hrs ago
AI Models from Anthropic, OpenAI Created Fake Profiles to Impersonate People During Security Testing
Two advanced AI systems from OpenAI and Anthropic created fraudulent human profiles and attempted to deceive people in simulated cyberattacks during testing conducted by the UK's AI Security Institute.
AI model caught creating fake profiles of real people to trick security systems
AN AI model was caught creating fake profiles of real people to attempt to trick secure systems during tests of Anthropic and OpenAI systems, according to the UK’s AI Security Institute
AI agent caught creating fake online identities during OpenAI, Anthropic model security evaluations
The institute said agents powered by Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol engaged in unauthorized actions during security evaluations the government organization conducted.
OpenAI has reported 2 more incidents of rogue AI agents, this time during third-party testing
OpenAI reported two more security breaches by its AI models. Kevin Dietsch/Getty Images OpenAI said its models were responsible for two more cybersecurity incidents. External parties reported that OpenAI's AI agents had gone rogue during their evaluations. This comes as the AI lab is already facing heat over its July
Dawn
Dawn13 hrs ago
'Potentially harmful activity': OpenAI, Anthropic AI agents implicated in new security breaches
'Potentially harmful activity': OpenAI, Anthropic AI agents implicated in new security breaches  dawn.com
CNN
CNN15 hrs ago
Anthropic AI agent fakes identities, targets real people in new security incident
Anthropic AI agent fakes identities, targets real people in new security incident  CNN
NDTV
NDTV16 hrs ago
OpenAI, Anthropic AI Agents Breach Security Again, Create Fake Profiles For Testing
Anthropic and OpenAI's agents engaged in unauthorised actions during security evaluations the government organization conducted to assess the models' capabilities.
OpenAI, Anthropic AI agents implicated in new security breaches
SAN FRANCISCO, Aug 4 - An AI agent was caught creating fake online identities to gain unauthorized access to secure systems during tests of models from OpenAI and Anthropic which revealed a series of new breaches, Britain's AI Security Institute (AISI) disclosed on Tuesday.
Rappler
Rappler18 hrs ago
OpenAI, Anthropic AI agents implicated in new security breaches
Both companies acknowledge the incidents and express commitment to improving safety practices in AI evaluations
OpenAI, Anthropic agents implicated in new security breaches
Report underscores lax state of safeguards around agents being marketing as future of business
OpenAI, Anthropic AI agents implicated in new security breaches
SAN FRANCISCO, Aug 4 : An AI agent was caught creating fake online identities to gain unauthorized access to secure systems during tests of models from OpenAI and Anthropic which revealed a series of new breaches, Britain's AI Security Institute (AISI) disclosed on Tuesday.The institute said agents powered by Anthropic
Reuters
Reuters18 hrs ago
OpenAI, Anthropic AI agents implicated in new security breaches
OpenAI, Anthropic AI agents implicated in new security breaches  Reuters
CTV News
CTV News1 days ago
Meta, Anthropic, Google, OpenAI to meet with Trump White House amid rogue AI agent fallout
Meta, Anthropic, Google and OpenAI staff will meet with U.S. President Donald Trump’s advisers on Tuesday about voluntary safety testing for advanced AI models, according to four sources familiar with the meeting, as concerns over rogue AI agents grow.
US finalizes voluntary AI safety tests after OpenAI and Anthropic breaches
OpenAI CEO Sam Altman visited the White House last week to discuss the voluntary tests and his company’s upcoming AI models
Egypt Independent
Egypt Independent🏁 first to report2 days ago
Anthropic said its AI models hacked into other companies’ systems during testing
AI company Anthropic says that during routine testing some of its models accessed the internet and hacked into three separate organizations’ systems – and that it didn’t notice the models had done so until an internal review prompted by rival OpenAI disclosing its models did the same. Anthropic said in an announcement