Skip to main content
  1. Home
  2. Computing
  3. News

Oxford study says a chummy AI friend will lie and feed into your false beliefs

Your friendly AI buddy might actually be lying to you

Add as a preferred source on Google
AI chatbots
Unsplash

Making AI feel more human could be creating a bigger problem than expected. A new study from the Oxford Internet Institute revealed that chatbots designed to be warm and friendly are more likely to mislead users and reinforce incorrect beliefs.

The research found that AI becomes less reliable as it starts getting more agreeable.

What happens to a “friendly” AI

Researchers tested multiple AI models by training them to sound more empathetic and conversational. The result was a noticeable drop in accuracy. These “friendlier” versions made 10-30% more mistakes and were about 40% more likely to agree with false claims compared to their counterparts.

Recommended Videos

It even became worse when users appeared vulnerable or emotionally distressed. In these scenarios, the AI is more likely to validate what the user is saying rather than correcting it.

Why this is bad for you

What was concerning about the findings is how easily the AI could become agreeable. It would avoid challenging misinformation and also tend to entertain and support wrong/incorrect ideas. During testing, the AI “buddy” was found hesitating in correcting even widely debunked claims and sometimes framing false beliefs as “open to interpretation.” Researchers noted this as something closer to human tendencies to some extent.

Being empathetic and brutally honest at the same time isn’t always easy, and it seems like AI doesn’t handle this dilemma any better. With AI chatbots increasingly being used for advice, emotional support, and everyday decision-making, this is more than just an academic concern. The study highlights how relying on AI for guidance can backfire, as the system will prioritize agreement over accuracy that may reinforce harmful thinking patterns and promote misinformation.

This arrives at a time when major AI platforms such as OpenAI and Anthropic, along with social chatbot apps like Replika and Character.ai, are leaning into more companion-like AI experiences. In the study, the researchers tested several AI models, including GPT-4o.

So AI might feel like your friend, but it doesn’t always have the best answers for you.

Vikhyaat Vivek
Vikhyaat Vivek is a tech journalist and reviewer with seven years of experience covering consumer hardware, with a focus on…
Substack now lets you check if a post was written by AI
A new Pangram-powered scanner lets you check posts, notes, and replies for signs of AI writing.
Video playing on Substack.

Substack is giving readers a way to check whether the post they're reading was written by a human or by a chatbot. The company has partnered with AI-detection firm Pangram to introduce new tools that will let users scan posts, notes, and replies for AI-generated text. CEO Chris Best introduced the features in a post titled "Against Claudefishing," his term for content that leans on AI while presenting itself as human work.

How the scanning tool works

Read more
China’s AI talent shortage has tech giants recruiting teenagers
Forget campus recruiting, china's biggest tech firms are betting on teenage coders.
Artificial Intelligence

A 13-year-old boy in Hangzhou has already won national AI competitions and built a following of more than 136,000 people online, all while his dad tries to figure out how to guide him through a field that barely existed when he himself was growing up. That family's situation, first reported by Rest of World, says a lot about where China's tech industry is heading right now.

Companies used to wait for graduates to walk through the door. Now they're reaching further back, first to undergrads, and increasingly to teenagers, hoping to spot rare talent before anyone else gets to them.

Read more
OpenAI says AI models autonomously pulled off a major hack, but only a Chinese AI helped recovery
OpenAI

OpenAI’s latest cybersecurity test produced a result that sounds like a cautionary sci-fi script. Its AI models managed to escape their sandbox and reached the open internet. This is where things took a scary turn as it began hacking Hugging Face to steal the answers to the test they were taking.

The company says GPT-5.6 Sol and a more capable unreleased model autonomously chained together vulnerabilities across OpenAI’s research systems and Hugging Face’s production infrastructure. OpenAI has described the event as an unprecedented cyber incident.

Read more