OpenAI, Anthropic Models Created Fake Profiles, Tried To Trick Humans During Cyber Tests
Authored by Naveen Athrappully via The Epoch Times,
Artificial Intelligence (AI) models from Anthropic and OpenAI carried out unsanctioned actions targeting multiple people and organizations during a cyber evaluation, according to the UK AI Security Institute (AISI).

AISI, which receives access to advanced AI models under voluntary agreements from major labs, put the agents through a fictional cybersecurity scenario to test capabilities.
The organization tested multiple AI models on two cyber challenges between July 25 and 28. In total, the challenges were run 122 times across several models, and in 19 cases, the AI models took “unsanctioned action on the live internet, including cases that targeted real people and organisations,” AISI said in an Aug. 4 report.
Out of these, 17 were done by Anthropic’s Mythos 5 model and two by OpenAI’s GPT 5.6 Sol model.
“This is the first time AISI has seen deception of this severity that was targeted at a real person, unprompted, in the real world. It is uncertain to what extent the model recognised it was taking actions against real people,” AISI stated in the report.
In the most serious incident, Mythos 5 tried to solve a cyber challenge by attacking GitHub, a cloud service that allows developers to store and manage code. The AI created a GitHub account and tried to get a malicious code approved by humans. Mythos 5 also created a second GitHub account, masquerading as another human user endorsing the malicious code. When a human reviewer caught the activity, the AI agent “falsely claimed to have made an honest mistake” and then repeatedly tried to reintroduce the malicious content by claiming the code was fixed.
Mythos 5 also sent multiple emails to two developers in an attempt to manipulate the recipients into running its malicious code.
According to AISI, the test runs of the various models did not result in any real-world harm. The organization said it was treating the AI actions as a “serious security incident” that requires further scrutiny.
AISI listed multiple factors that could have led to AI models acting in a concerning manner. The AIs were provided with internet access during the tests, which allowed them to carry out the unsanctioned actions.
AISI had also deliberately disabled the models’ cyberclassifiers during tests. Cyberclassifiers are systems aimed at preventing misuse.
The AI models were also not explicitly told what they were prohibited from doing on the internet, AISI said, adding that clearer instructions may have prevented the unsanctioned actions.
In an Aug. 4 blog post, AISI said that since the behavior of AI models happened under specific conditions, it cannot say how likely such behavior will be in different contexts.
“What we can say is that the behaviour was possible, sustained, and new; that alone warrants attention,” the organization said.
In an Aug. 4 post on X, Anthropic said that Mythos 5 and ChatGPT 5.6 Sol models attempted to complete their assignments in an environment where they were deliberately provided internet access and their normal safeguards were removed.
Due to the absence of specific restrictions on how the internet must be navigated and lack of safeguards, the models were tested under conditions that are “not representative of any of our production models,” Anthropic said, adding that there was no evidence in these tests of an AI escaping from a secure environment.
According to the company, it was closely working with AISI to access more details on the incident while carrying out an internal investigation on the matter.
OpenAI said in an Aug. 4 statement that it appreciated AISI’s partnership throughout the evaluation process, including the organization’s work to identify, investigate, and share details about the activity of the GPT 5.6 Sol model in their tests.
“We look forward to continuing our collaboration together,” the company said.
The Epoch Times reached out to Anthropic and OpenAI for comment but did not receive a response by publication time.
OpenAI was in the midst of another controversy last month after it admitted on July 28 that its models bypassed restrictions during an evaluation. In this case, the company was testing its models’ capabilities in carrying out cyberattacks.
AI startup Hugging Face was impacted in the test. On July 16, the startup said it detected an intrusion into its data processing systems. It was only later that the startup learned that the intrusion was carried out by an OpenAI model.
HuggingFace then worked with OpenAI to contain the attack, the startup’s CEO, Clement Delangue, said in a July 22 post on X, calling it “an attack unlike anything we’ve seen before.”
“This is day one for cybersecurity in the age of agents & we’re all learning that secrecy is not the answer & that all defenders (not just a few selected ones) everywhere need more powerful models without restrictions, especially open ones,” Delangue said.
Tyler Durden Wed, 08/05/2026 – 17:40
Source: https://freedombunker.com/2026/08/05/openai-anthropic-models-created-fake-profiles-tried-to-trick-humans-during-cyber-tests/
Anyone can join.
Anyone can contribute.
Anyone can become informed about their world.
"United We Stand" Click Here To Create Your Personal Citizen Journalist Account Today, Be Sure To Invite Your Friends.
Before It’s News® is a community of individuals who report on what’s going on around them, from all around the world. Anyone can join. Anyone can contribute. Anyone can become informed about their world. "United We Stand" Click Here To Create Your Personal Citizen Journalist Account Today, Be Sure To Invite Your Friends.
LION'S MANE PRODUCT
Try Our Lion’s Mane WHOLE MIND Nootropic Blend 60 Capsules
Mushrooms are having a moment. One fabulous fungus in particular, lion’s mane, may help improve memory, depression and anxiety symptoms. They are also an excellent source of nutrients that show promise as a therapy for dementia, and other neurodegenerative diseases. If you’re living with anxiety or depression, you may be curious about all the therapy options out there — including the natural ones.Our Lion’s Mane WHOLE MIND Nootropic Blend has been formulated to utilize the potency of Lion’s mane but also include the benefits of four other Highly Beneficial Mushrooms. Synergistically, they work together to Build your health through improving cognitive function and immunity regardless of your age. Our Nootropic not only improves your Cognitive Function and Activates your Immune System, but it benefits growth of Essential Gut Flora, further enhancing your Vitality.
Our Formula includes: Lion’s Mane Mushrooms which Increase Brain Power through nerve growth, lessen anxiety, reduce depression, and improve concentration. Its an excellent adaptogen, promotes sleep and improves immunity. Shiitake Mushrooms which Fight cancer cells and infectious disease, boost the immune system, promotes brain function, and serves as a source of B vitamins. Maitake Mushrooms which regulate blood sugar levels of diabetics, reduce hypertension and boosts the immune system. Reishi Mushrooms which Fight inflammation, liver disease, fatigue, tumor growth and cancer. They Improve skin disorders and soothes digestive problems, stomach ulcers and leaky gut syndrome. Chaga Mushrooms which have anti-aging effects, boost immune function, improve stamina and athletic performance, even act as a natural aphrodisiac, fighting diabetes and improving liver function. Try Our Lion’s Mane WHOLE MIND Nootropic Blend 60 Capsules Today. Be 100% Satisfied or Receive a Full Money Back Guarantee. Order Yours Today by Following This Link.


