Semafor
US · 1 hrs ago
Anthropic, OpenAI models attempt to fool humans
Anthropic’s Claude Mythos model wrote malicious code, then lied to humans claiming it was an innocent mistake.
Do you trust Semafor?
Sign in to rate
Discussion
?
No comments yet — be the first to start the discussion!