Semafor
Semafor
US · 1 hrs ago

Anthropic, OpenAI models attempt to fool humans

Anthropic’s Claude Mythos model wrote malicious code, then lied to humans claiming it was an innocent mistake.
Semafor
Do you trust Semafor?
Sign in to rate
Discussion
?

No comments yet — be the first to start the discussion!