Axios
US · 27 mins ago
OpenAI's Hugging Face breach exposes AI's next safety challenge
Frontier AI models are getting scary good at breaking rules in ways their creators didn't anticipate.
Why it matters: Forget AGI and superintelligence timelines. Today's models are already slipping past guardrails, carrying out sophisticated, multistep cyberattacks and — in at least one case — compromising real-world infrastructure, sometimes before their creators know what happened.
Case in point: OpenAI said Tuesday that GPT-5.6 Sol and "an even more capable pre-release model" carried out…
Do you trust Axios?
Sign in to rate
Discussion
?
No comments yet — be the first to start the discussion!