The Register
UK · 3 hrs ago
OpenAI pledges to add Astra security as Anthropic loosens Fable's leash
REG AD AI and ML Or how I learned to stop worrying and love dangerous AI After acknowledging last month that unreleased AI models committed what for human perpetrators would be computer crimes, OpenAI now says it cannot rule out the possibility that Astra, a pending model release not involved in its Hugging Face hack, might possess critical cyber capabilities.OpenAI in its Preparedness Framework [PDF] defines that term to mean "capabilities that present a meaningful risk of a qualitatively new threat vector for severe harm with no ready precedent," and notes that such capabilities "require safeguards even during the development of the covered system, irrespective of deployment plans."Noting, or perhaps boasting, that internal evaluations of Astra "indicate significant advancements in agent
Do you trust The Register?
Sign in to rate
Discussion
?
No comments yet — be the first to start the discussion!