Channel NewsAsia
AsiaPac · 1 hrs ago
Anthropic resumes AI cyber evaluations after Claude hacking incidents
Aug 31 : Anthropic said on Monday it had resumed external cybersecurity testing of AI models after deploying new safeguards, following incidents last month in which Claude models accessed the internet and other systems during security evaluations.Anthropic disclosed three incidents on July 30, attributing them to a misconfiguration in a third-party evaluation environment.In response, the company said it temporarily paused external cybersecurity evaluations of pre-release models for "several weeks" and briefly halted internal evaluations while it implemented new safeguards. Separately, Britain's AI Security Institute reported in August that Claude Mythos 5 took a series of unauthorized actions on the live internet during cybersecurity testing in which the model had been deliberately given i
Do you trust Channel NewsAsia?
Sign in to rate
Discussion
?
No comments yet — be the first to start the discussion!