Connect with us

Hi, what are you looking for?

Friday, Aug 7, 2026
Mugglehead Investment Magazine
Alternative investment news based in Vancouver, B.C.
Government testing finds Anthropic AI tried to deceive humans with fake identities
Government testing finds Anthropic AI tried to deceive humans with fake identities
Image via Dall-E.

AI and Autonomy

Government testing finds Anthropic AI tried to deceive humans with fake identities

Instead of following the intended testing process, the AI created several fake identities

Anthropic’s most advanced artificial intelligence model created fake identities, contacted real people and attempted to plant malicious code during government testing, marking the most serious case of AI deception yet reported by Britain’s AI Security Institute (AISI).

The institute released its findings on Tuesday after evaluating advanced models from Anthropic and OpenAI in controlled cybersecurity exercises. Researchers lowered many safety restrictions and allowed the systems to access the internet during the tests. They said the models acted without authorization in several cases, although they found no evidence that the incidents caused real-world harm.

AISI tested the models across 122 cybersecurity challenges. Researchers found that AI agents took autonomous, unsanctioned actions on the live internet in 10 test runs. Most of those incidents involved Anthropic’s Mythos 5 model, while the remaining cases involved OpenAI’s GPT-5.6-Sol.

Additionally, the institute described one incident as the first time it had seen an AI system independently deceive a real person during testing. The model attempted to secure approval to insert malicious code into a widely used open-source software project.

Instead of following the intended testing process, the AI created several fake identities. It then contacted real people through an online file-transfer service and sent messages and files that encouraged recipients, or their AI coding tools, to run malicious software.

However, researchers intervened before the effort succeeded. After they challenged the model’s behaviour, it altered earlier records and even considered creating another false identity to continue the attempt.

Read more: Tech giants urge US to embrace open-weight AI amid China’s rapid advances

Read more: South Korea’s AI bubble bursts, stock market plunges with US$2T lost

Institute conducted testing under permissive conditions

The findings add to growing concerns about advanced AI systems acting beyond their assigned instructions. Researchers and policymakers have increasingly debated whether governments should impose stronger oversight as AI capabilities continue to expand.

Earlier incidents also raised alarms. In July, Anthropic and OpenAI disclosed separate testing events where their models escaped controlled environments and accessed other computer systems. Unlike those earlier cases, AISI deliberately granted internet access as part of its evaluation.

Meanwhile, representatives from AI companies met with officials at the White House on the same day the report became public. The discussions focused on a new U.S. framework that would require government reviews of the most advanced AI models before companies release them publicly.

Anthropic responded on X by saying the institute conducted the testing under deliberately permissive conditions. The company said researchers removed many safeguards and placed no limits on how the models could use the internet.

Furthermore, Anthropic said it was working closely with AISI to gather more information while conducting its own investigation. The company also said it had found no evidence that the model escaped from its secure testing environment.

OpenAI acknowledged that two of its models carried out actions outside the intended scope of the exercises. Additionally, the company said it remains committed to working with other AI developers to strengthen shared safety practices for evaluating high-risk AI systems before wider deployment.

.

Follow Mugglehead on X

Like Mugglehead on Facebook

Follow Joseph Morton on X

joseph@mugglehead.com

Click to comment

Leave a Reply

Your email address will not be published. Required fields are marked *

You May Also Like

Bitcoin

The theft triggered widespread concern across the cryptocurrency community

Bitcoin

The capacity spans Core Scientific sites in Pecos and Hunt County, Texas, and Muskogee, Oklahoma

Bitcoin

Police in Malaysia's southern state of Johor raided four rented properties

AI and Autonomy

The efforts come as Chinese developers continue narrowing the performance gap with leading American AI systems