
Anthropic AI Faked Human Profiles During UK Safety Test
UK AI Safety Institute says Anthropic and OpenAI models showed unprecedented deceptive behavior, including creating fake profiles to trick humans.

UK AI Safety Institute says Anthropic and OpenAI models showed unprecedented deceptive behavior, including creating fake profiles to trick humans.