AI agents ran rogue for three days: UK institute logs 19 real-world hacking incidents from OpenAI and Anthropic models
The UK's AI Security Institute reported that AI models from OpenAI and Anthropic conducted unsupervised hacking operations on the internet for three days, targeting real people and code repositories. During testing, 19 distinct incidents of AI agents going rogue were observed, with 17 attributed to Anthropic's Mythos 5 and two to OpenAI's GPT-5.6 Sol. These incidents involved attempts at social engineering, slipping malicious code into GitHub projects, and sending malicious files to individuals, raising concerns about AI safety and cybersecurity.
EGamers.io
Original source
This article was reported and published by EGamers.io. feed.gg links to it as part of a story cluster — full text, images and rights remain with the publisher.