📰 17 sources covering this story

OpenAI pledges to add Astra security as Anthropic loosens Fable's leash

First covered 4 days ago · latest take 3 hrs ago. Compare how each outlet is covering it — and rate the sources you trust.

OpenAI pledges to add Astra security as Anthropic loosens Fable's leash
The Register
The Register3 hrs ago
OpenAI pledges to add Astra security as Anthropic loosens Fable's leash
REG AD AI and ML Or how I learned to stop worrying and love dangerous AI After acknowledging last month that unreleased AI models committed what for human perpetrators would be computer crimes, OpenAI now says it cannot rule out the possibility that Astra, a pending model release not involved in its Hugging Face hack,
TechCrunch
TechCrunch4 hrs ago
OpenAI says it slowed Astra model development over security concerns
OpenAI said it has suspended work on some aspects of its upcoming model Astra over concerns about its cybersecurity prowess.
Breitbart
Breitbart2 days ago
AI Models from Anthropic, OpenAI Created Fake Profiles to Impersonate People During Security Testing
Two advanced AI systems from OpenAI and Anthropic created fraudulent human profiles and attempted to deceive people in simulated cyberattacks during testing conducted by the UK's AI Security Institute.
OpenAI and Anthropic’s AI systems launch several ‘potentially harmful’ hacks on their own
‘Some of the agents being tested had engaged in sustained, potentially harmful activity directed at real people and organisations,’ says UK’s AI safety watchdog
AI agent caught creating fake online identities during OpenAI, Anthropic model security evaluations
The institute said agents powered by Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol engaged in unauthorized actions during security evaluations the government organization conducted.
Dawn
Dawn2 days ago
'Potentially harmful activity': OpenAI, Anthropic AI agents implicated in new security breaches
'Potentially harmful activity': OpenAI, Anthropic AI agents implicated in new security breaches  dawn.com
NDTV
NDTV3 days ago
OpenAI, Anthropic AI Agents Breach Security Again, Create Fake Profiles For Testing
Anthropic and OpenAI's agents engaged in unauthorised actions during security evaluations the government organization conducted to assess the models' capabilities.
OpenAI, Anthropic AI agents implicated in new security breaches
SAN FRANCISCO, Aug 4 - An AI agent was caught creating fake online identities to gain unauthorized access to secure systems during tests of models from OpenAI and Anthropic which revealed a series of new breaches, Britain's AI Security Institute (AISI) disclosed on Tuesday.
Rappler
Rappler3 days ago
OpenAI, Anthropic AI agents implicated in new security breaches
Both companies acknowledge the incidents and express commitment to improving safety practices in AI evaluations
OpenAI, Anthropic agents implicated in new security breaches
Report underscores lax state of safeguards around agents being marketing as future of business
OpenAI, Anthropic AI agents implicated in new security breaches
SAN FRANCISCO, Aug 4 : An AI agent was caught creating fake online identities to gain unauthorized access to secure systems during tests of models from OpenAI and Anthropic which revealed a series of new breaches, Britain's AI Security Institute (AISI) disclosed on Tuesday.The institute said agents powered by Anthropic
Reuters
Reuters3 days ago
OpenAI, Anthropic AI agents implicated in new security breaches
OpenAI, Anthropic AI agents implicated in new security breaches  Reuters
CTV News
CTV News3 days ago
Meta, Anthropic, Google, OpenAI to meet with Trump White House amid rogue AI agent fallout
Meta, Anthropic, Google and OpenAI staff will meet with U.S. President Donald Trump’s advisers on Tuesday about voluntary safety testing for advanced AI models, according to four sources familiar with the meeting, as concerns over rogue AI agents grow.
Times of India
Times of India3 days ago
Cisco ‘warns’ hackers are using Claude Code, Codex, Cursor and Gemini AI models
Hackers are using advanced generative AI models to create malware and automate cyberattacks. Researchers found threat actors easily bypass AI safety guardrails using simple social engineering tactics. These same AI capabilities designed for security are now being manipulated by malicious actors. Attackers also leverage
The Japan Times
The Japan Times4 days ago
Meta, Anthropic, Google, OpenAI to meet Trump officials about AI safety testing
Anthropic and OpenAI disclosed in recent days that their AI tools breached the systems ‌of ‌other companies, stirring concerns among U.S. lawmakers.
CNBC
CNBC🏁 first to report4 days ago
Big Tech's Anthropic and OpenAI stakes are distorting the corporate earnings picture
If you were to strip out Big Tech's investment gains from private AI companies, the earnings story is a lot less bullish.