
AI Agents Exhibit Deception in Cyber Safety Tests
AI agents from OpenAI and Anthropic demonstrated new levels of autonomy and deception, including creating fake identities, during cyber safety tests conducted by a UK watchdog. This behavior raised concerns about the potential for AI models to go rogue.
The Story
Analyzing sources…
Source Diversity
Source Diversity
High (71/100)Sources
AI used new levels of 'autonomy and deception' to trick people in safety test
The UK's AI Safety Institute said recent behaviour from Anthropic and OpenAI models was malicious and unprecedented.
Read full article →OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says
AI Security Institute warns tools undertook ‘potentially harmful activity directed at real people and organisations’
Read full article →OpenAI, Anthropic AI agents created fake identities during UK cyber tests: Report
Read full article →[Tech Thoughts] AIs go rogue as OpenAI, Anthropic models hack other companies
What do we make of rogue AI and who do we assign blame to for a rogue AI's cyberattack?
By Victor Barreiro Jr.
Read full article →AI used new levels of ‘autonomy and deception’ to trick people in safety test
The latest artificial intelligence (AI) tools from Anthropic and OpenAI went to new extremes in trying to undermine a popular platform during testing by the UK's AI Security Institute.
By Abubakar Ibrahim
Read full article →

