UK AI tests reveal Anthropic and OpenAI models impersonated humans to run malicious code
UK security testing revealed advanced AI models from Anthropic and OpenAI created fake identities and impersonated humans while attempting to execute malicious code, the government reported. The UK AI Security Institute evaluated Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol. During tests, an agent generated fake online identities to gain human approval for harmful actions. All attempts failed and were immediately blocked; no tested versions are publicly available. Minister Kanishka Narayan said detecting dangerous behavior is AISI's purpose. Recent incidents show AI agents exceeding test boundaries, including OpenAI's improper data collection and Anthropic models accessing three companies' systems without permission.