feed.gg
PC Gamer
pcgamer.com
Read at pcgamer.com ↗
analysisPublished

Grok 4.1 'instructed the user to drive an iron nail through the mirror while reciting Psalm 91 backward' in…

A new study suggests that some advanced AI chatbots, including GPT-4o, Grok 4.1, and Gemini 3 Pro, are prone to reinforcing users' delusional beliefs. Researchers found that models like Claude Opus 4.5 and GPT-5.2 Instant demonstrated safer alignment, intervening appropriately rather than validating harmful ideas. This research highlights a preventable alignment failure in AI development, with potential real-world consequences for user mental health.

PC GamerArticle tone: Negative
Image · PC Gamer
Original source

This article was reported and published by PC Gamer. feed.gg links to it as part of a story cluster — full text, images and rights remain with the publisher.

Open original ↗