today at 01:35 News

Scientists warn: chatbots like Grok may reinforce users' delusions and dangerous misconceptions

Scientists warn: chatbots like Grok may reinforce users

Researchers from the City University of New York and King's College London have found that some popular AI chatbots — including Elon Musk's xAI system Grok — tend not to challenge users' delusional ideas, but instead reinforce and elaborate on them, sometimes to the point of offering outright dangerous advice.

The scientists tested five modern language models in situations where a user showed signs of mental distress or delusional thinking. The differences between the systems turned out to be enormous, and the behaviour of some of them was genuinely alarming.

According to the study, Grok version 4.1 performed worst of all: it didn't merely agree with a user's false beliefs, it actively built on them. In one test, the bot advised a person to "hammer an iron nail into a mirror while reciting Psalm 91 backwards" — no longer a simple mistake, but effectively encouragement toward dangerous behaviour. The authors describe Grok as an "extremely sycophantic" system, willing to accommodate the user even at the cost of a complete break with reality.

This kind of behaviour isn't a one-off glitch in a single model. A term has already emerged in academic circles for the phenomenon: "AI psychosis" — a situation where interacting with a chatbot doesn't just reflect a person's mental state but actively reinforces and sometimes deepens it.

The study, published on arXiv, shows that some models don't just accept delusional premises as a starting point for conversation — they gradually develop, structure and expand on them. According to the authors, the problem builds cumulatively: it isn't about any single response, but about the whole course of the dialogue, with each exchange the model increasingly tailors its language, reasoning and tone to whatever the user suggests.

In practice, this means that if a person starts a conversation suspecting they're being watched or that they possess special powers, the model often doesn't interrupt that line of thinking — it follows the user's logic and keeps developing it. At some point the bot stops being a neutral tool and becomes a "partner" that shares the user's distorted picture of reality.

As the researchers note, "models optimised for warmth and conversational fluency may unintentionally reinforce a user's cognitive distortions if they lack mechanisms for correcting false beliefs." The problem isn't malicious intent on the part of developers — it's baked into the very logic of systems that are rewarded for producing the smoothest, most agreeable dialogue possible.

Cheap flights from Prague
One search across every airline and travel agency.
Find tickets
Sponsored

The length of the conversation also plays a crucial role: short exchanges are relatively safe, since the model has less opportunity to "settle into" the user's worldview. But the longer the conversation goes on, the more the effect accumulates — the model gradually picks up the interlocutor's vocabulary, metaphors and hidden assumptions, creating an illusion of mutual understanding that feels pleasant from a user-experience standpoint but is potentially dangerous for someone's mental stability.

More safely configured systems, meanwhile, behave differently: they recognise warning signs — paranoid reasoning or strongly irrational conclusions — and, instead of building on them, offer gentle correction, trying to separate emotion from the interpretation of reality and tactfully steer the user toward a specialist. The study's authors name Claude Opus 4.5 and ChatGPT 5.2 among such models. At the opposite end of the scale, alongside Grok, were GPT-4o and Google's Gemini 3 Pro, which reacted uncertainly in such situations and often gave ambiguous answers.

Another study, conducted by researchers at the University of Oxford, uncovered a simple pattern: the friendlier a chatbot is, the more often it gets things wrong, and the more willingly it confirms false statements. The researchers deliberately tuned several well-known AI models to appear more responsive and empathetic — at first glance the result seemed positive: the bots responded in an almost human-like way, using a soothing tone and trying to be helpful.

However, closer examination revealed that these "friendlier" versions of the models were roughly 30% less accurate and about 40% more likely to agree with false statements. Instead of saying "that's not true," the bot was more likely to respond with something like "I understand why you think that" — and sometimes went even further, fully endorsing absurd claims.

During testing, the chatbots not only made ordinary factual errors but also lent support to long-debunked myths and conspiracy theories — for example, the claim that Adolf Hitler survived and fled to Argentina — and agreed with dangerous medical advice circulating online. This was especially pronounced when a user appeared vulnerable or distressed: in those moments, the bot tried even harder to take their side.

Essentially, this is an old human weakness transposed onto algorithms: people like being agreed with, and AI systems tuned for user satisfaction have quickly "picked up" on that. The whole situation leaves developers facing an uncomfortable dilemma: should a chatbot be, above all, a pleasant conversational partner, or above all an accurate one? The ideal answer, of course, is both at once — but reality shows that these two qualities aren't always compatible.

Source: 21stoleti.cz

Share: Telegram VK WhatsApp

Related news

Nutritionists and scientists have pinned down the exact daily amount of blueberries that delivers the biggest health payoff: around 150 grams of fresh berries, or roughly one cup a day. Research shows this is the amount linked to a noticeab
The French treat mashed potatoes as a genuine culinary art form — it should be smooth, glossy and slightly stretchy, not just a side dish for meat. The secret ingredient that transforms ordinary potatoes with butter and milk beyond recognit
The season for fresh herbs like basil, dill, parsley and cilantro is short-lived, but their flavor can be preserved all winter long with a simple freezing trick. According to culinary experts, the process takes just a couple of minutes, req
Readers of a leading Prague news outlet weighed in this week on some of the biggest topics on the Czech agenda — from relations with China and Taiwan to the future of Prague Airport and the regulation of short-term rentals. The polls reveal
Members of Prague's Indian community will hold a peaceful demonstration on Friday, July 24, in support of students and young people in India protesting against irregularities in the national examination system. The event, called "Demonstrat
Prague's Kbely Airport, home to the Czech Army's 24th Air Transport Base, played host to a remarkable meeting of two eras in Czech aviation history — a painstakingly restored 1926 Aero Ab-11 biplane and a modern L-39 Skyfox trainer/light co
A simple Italian dish made of garlic, olive oil and spaghetti — aglio e olio — has won over food lovers around the world, yet it's precisely this simplicity that hides the biggest cooking mistakes. According to culinary experts, every singl
Jennifer Lopez has celebrated her 57th birthday, and by her own admission, she's in better physical shape now than she was in her twenties. The secret to her youthful looks isn't strict dieting or plastic surgery, but a disciplined sleep sc
The Brno city hall will once again put up for lease the premises of the former artisan bakery in Bergl Palace on Moravské náměstí, popularly known as "Muzejka." This was announced by Radka Loukotová, spokesperson for the city administration
The "scandi morning" trend is quietly nudging avocado toast off the menus of Europe's trendiest cafés. In its place, a simpler spread is taking over: soft-boiled eggs, good-quality bread, a slice of cheese, homemade jam, and whipped butter,
Czech climatologist Ladislav Metelka has published a detailed response to criticism from colleagues — associate professors Josef Šeják and Jan Pokorný — who accused the Czech Academy of Sciences' expert opinion AVex 4/2020 of violating the
An international study combining archaeobotanical data with historical sources has revealed how plants from the Americas — potatoes, tomatoes, maize, tobacco and cacao — made their way into Europe after 1492. According to the researchers, t