The Register
The Register
UK · 3 hrs ago

OpenAI pledges to add Astra security as Anthropic loosens Fable's leash

REG AD AI and ML Or how I learned to stop worrying and love dangerous AI After acknowledging last month that unreleased AI models committed what for human perpetrators would be computer crimes, OpenAI now says it cannot rule out the possibility that Astra, a pending model release not involved in its Hugging Face hack, might possess critical cyber capabilities.OpenAI in its Preparedness Framework [PDF] defines that term to mean "capabilities that present a meaningful risk of a qualitatively new threat vector for severe harm with no ready precedent," and notes that such capabilities "require safeguards even during the development of the covered system, irrespective of deployment plans."Noting, or perhaps boasting, that internal evaluations of Astra "indicate significant advancements in agent
The Register
Do you trust The Register?
Sign in to rate
Discussion
?

No comments yet — be the first to start the discussion!