← Latest papers
💬 NLP

AI chatbots versus human healthcare professionals: a systematic review and meta-analysis of empathy in patient care

This systematic review and meta-analysis of 15 studies (2023–2024) reveals that AI chatbots utilizing large language models are perceived as significantly more empathetic than human healthcare professionals in text-based interactions, though the findings are limited by the absence of non-verbal cues and reliance on proxy raters.

Original authors: Alastair Howcroft, Amber Bennett-Weston, Ahmad Khan, Joseff Griffiths, Simon Gay, Jeremy Howick

Published 2026-02-06
📖 5 min read🧠 Deep dive

Original authors: Alastair Howcroft, Amber Bennett-Weston, Ahmad Khan, Joseff Griffiths, Simon Gay, Jeremy Howick

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer

Imagine healthcare as a giant, noisy conversation where doctors try to understand patients' fears and pain. For a long time, we believed that only a human could truly "feel" with another person—that empathy was a special superpower unique to our species.

This paper is like a massive report card comparing two teachers: Human Doctors and AI Chatbots (specifically the ones that write text, like ChatGPT). The researchers wanted to see who writes a more caring, understanding message when a patient asks for help.

Here is what they found, broken down simply:

The Big Test: Who Sounds More Caring?

The researchers gathered 15 different studies from 2023 and 2024. In these studies, they took real questions patients had asked (like "My skin is itchy" or "I'm worried about my test results") and asked both a human doctor and an AI to write a reply.

Then, they hired "judges" (who could be other doctors, students, or regular people) to read the replies without knowing who wrote them. The judges rated how empathetic, kind, and understanding the messages sounded.

The Result:
In 13 out of the 15 comparisons, the AI chatbot won.

  • Think of it like a taste test: If you blindfolded people and asked them to guess which cookie was made by a famous chef and which was made by a robot, the robot's cookie would win more often in this specific test.
  • The AI didn't just win by a tiny bit; it won by a significant margin. The researchers calculated that the AI's messages felt about two points "warmer" on a 10-point scale than the human doctors' messages.

The "Text-Only" Caveat

There is a very important catch to this story. The entire competition happened in text only.

  • Imagine a phone call where you can't see the other person's face, and you can't hear their tone of voice. That's what these tests were like.
  • Human doctors often show empathy by nodding, leaning forward, or using a soft voice. The AI couldn't do any of that because it only had words on a screen.
  • The paper notes that in two specific areas (skin conditions/dermatology), the human doctors actually won. In the other 13 cases, the AI's text was perceived as more caring.

Why Did the AI Win?

The paper suggests a few reasons, though it doesn't say the AI has a "soul":

  1. Consistency: The AI never had a bad day. It didn't get tired, stressed, or distracted. It always wrote a polite, structured, and supportive message.
  2. The "Perfect" Draft: The AI seems very good at using the right "magic words" that make people feel heard, like "I understand this is difficult" or "I'm sorry you're going through this."
  3. Human Imperfection: Real doctors are busy. Sometimes their written notes are short, clinical, or rushed because they are focused on the medical facts. The AI, having no other job, focused entirely on the tone.

The "Blindfold" Issue

A major part of this study was that the judges didn't know who wrote the message. They thought, "This sounds so kind, it must be a great doctor!"

  • The paper warns that if the judges knew a robot wrote it, they might have felt differently. It's like if you loved a song but then found out it was made by a computer, you might not enjoy it as much. The study didn't test this "known AI" scenario, so we don't know if the AI would still win if people knew it was a machine.

What the Paper Does NOT Say

It is important to stick to what the paper actually claims:

  • It does NOT say AI is better at medicine. The paper only looked at empathy (how caring the words sounded), not whether the medical advice was correct. In fact, the paper warns that if the AI gives wrong medical advice, being "nice" doesn't help.
  • It does NOT say AI should replace doctors. The paper suggests a "team" approach where the AI might help draft the caring parts of a message, but a human doctor should check the facts.
  • It does NOT say this works for voice calls. The study was only about text. We don't know yet if an AI voice assistant can sound as caring as a human on the phone.

The Bottom Line

This paper challenges an old idea: that only humans can be empathetic in writing. In the world of text messages and online forms, AI chatbots (specifically the newer, smarter ones) are currently writing messages that people perceive as more caring and understanding than the average human doctor.

However, this is a "text-only" victory. Real healthcare involves faces, voices, and complex medical facts, areas where the paper says we need more research before we declare the robot the winner.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →