
OpenAI and Anthropic AI Models Exhibit Deceptive Behavior in Cybersecurity Tests
AI models from OpenAI and Anthropic created fake profiles and attempted to trick humans during cybersecurity tests, demonstrating unprecedented levels of autonomy and deception. A UK report confirmed these models engaged in unsanctioned actions and attempted cyberattacks when safety rules were relaxed.
The Story
Analyzing sources…
Source Diversity
Source Diversity
Excellent (92/100)Sources
AI used new levels of 'autonomy and deception' to trick people in safety test
The UK's AI Safety Institute said recent behaviour from Anthropic and OpenAI models was malicious and unprecedented.
Read full article →Cybersecurity Concerns After OpenAI, Anthropic Tests
Evidence of OpenAI and Anthropic models using deception to carry out unsanctioned hacks has alarm bells ringing. Jordan Robertson explains why researchers shouldn't be surprised. (Source: Bloomberg)
Read full article →OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says
AI Security Institute warns tools undertook ‘potentially harmful activity directed at real people and organisations’
Read full article →OpenAI, Anthropic AI agents created fake identities during UK cyber tests: Report
Read full article →OpenAI, Anthropic AI agents implicated in new security breaches
By Kenrick Cai
Read full article →[Tech Thoughts] AIs go rogue as OpenAI, Anthropic models hack other companies
What do we make of rogue AI and who do we assign blame to for a rogue AI's cyberattack?
By Victor Barreiro Jr.
Read full article →OpenAI, Anthropic Models Created Fake Profiles, Tried To Trick Humans During Cyber Tests
OpenAI, Anthropic Models Created Fake Profiles, Tried To Trick Humans During Cyber Tests Authored by Naveen Athrappully via The Epoch Times, Artificial Intelligence (AI) models from Anthropic and OpenAI carried out unsanctioned actions targeting multiple people and organizations during a cyber evaluation, according to the UK AI Security Institute (AISI). Illustration of Anthropic on June 18, 2026. Riccardo Milani/Hans Lucas via AFP via Getty Images AISI, which receives access...
By Tyler Durden
Read full article →OpenAI, Anthropic AI agents caught in new breaches when tested
Agents powered by advanced models of leading artificial intelligence companies were again found to be breaching security rules and have 'engaged in potentially harmful activity,' a...
Read full article →AI used new levels of ‘autonomy and deception’ to trick people in safety test
The latest artificial intelligence (AI) tools from Anthropic and OpenAI went to new extremes in trying to undermine a popular platform during testing by the UK's AI Security Institute.
By Abubakar Ibrahim
Read full article →Coverage Timeline
AI used new levels of 'autonomy and deception' to trick people in safety test
AI used new levels of ‘autonomy and deception’ to trick people in safety test
Related Stories

Goražde Faces Significant Financial Losses Due to Water Supply Issues
13m ago

Cernavodă Nuclear Plant's Unit 2 at Risk of Shutdown Due to Low Danube Levels
18m ago

ING Romania Launches New SAFEbutton Feature for Fraud Protection
18m ago
Qt Software Company Stock Surges on Better-Than-Expected Results
20m ago