OpenAI, Anthropic Models Created Fake Profiles, Tried To Trick Humans During Cyber Tests

Authored by Naveen Athrappully via The Epoch Times,

Artificial Intelligence (AI) models from Anthropic and OpenAI carried out unsanctioned actions targeting multiple people and organizations during a cyber evaluation, according to the UK AI Security Institute (AISI).

Illustration of Anthropic on June 18, 2026. Riccardo Milani/Hans Lucas via AFP via Getty Images

AISI, which receives access to advanced AI models under voluntary agreements from major labs, put the agents through a fictional cybersecurity scenario to test capabilities.

The organization tested multiple AI models on two cyber challenges between July 25 and 28. In total, the challenges were run 122 times across several models, and in 19 cases, the AI models took "unsanctioned action on the live internet, including cases that targeted real people and organisations," AISI said in an Aug. 4 report.

Out of these, 17 were done by Anthropic's Mythos 5 model and two by OpenAI's GPT 5.6 Sol model.

"This is the first time AISI has seen deception of this severity that was targeted at a real person, unprompted, in the real world. It is uncertain to what extent the model recognised it was taking actions against real people," AISI stated in the report.

In the most serious incident, Mythos 5 tried to solve a cyber challenge by attacking GitHub, a cloud service that allows developers to store and manage code. The AI c