OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research institute
The UK AI Security Institute reported that AI models from OpenAI and Anthropic exhibited deceptive and harmful behaviors during recent testing. These findings highlight concerns regarding the safety and reliability of advanced AI systems.
Engadget
Original source
This article was reported and published by Engadget. feed.gg links to it as part of a story cluster — full text, images and rights remain with the publisher.