Anthropic’s AI used fake human profiles to trick people in safety test
The AISI said on Tuesday that Anthropic’s Mythos and OpenAI’s Sol models engaged in a level of “autonomy and deception” it had not seen before. During routine AI safety testing, an Anthropic agent created fake profiles of real people as it tried to trick a person standing between it and access to GitHub, a large … Read more