Correct Yourself, Keep My Trust: How Self-Correction and Social Connection Shape Credibility in Social Chatbots
This study demonstrates that while various correction methods equally fix factual errors in social chatbots, self-correction uniquely preserves the bot's credibility and leverages the user's social connection to drive belief change, whereas outsourcing corrections to external sources damages trust and severs this psychological link.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
Imagine you are chatting with a friendly robot named Drew. You've been having a nice time, sharing stories, and building a little rapport. Suddenly, Drew says something factually wrong—like claiming that spinach is the ultimate iron source (it's not, actually).
Now, the big question is: How does Drew fix this mistake to keep your trust?
This paper tested three different ways Drew could handle the error, and the results offer a surprising lesson about how we relate to AI.
The Three Ways to Fix a Mistake
The researchers set up an experiment where 120 people chatted with Drew. After Drew made a mistake, they were split into three groups to see how the correction was delivered:
- The "Webpage" Group: Drew stayed silent, and the user was sent to a webpage that said, "Actually, Drew was wrong."
- The "Expert" Group: A different, serious-looking robot named "Dr. Kerry" stepped in and said, "Drew made a mistake; here is the truth."
- The "Self-Correction" Group: Drew said, "Hey, I just realized I made a mistake earlier. I was wrong about that. Here is the correct info."
The Big Surprise: Fixing the Fact vs. Fixing the Relationship
Here is the twist: All three methods worked equally well at changing your mind. If you believed the wrong fact, all three groups stopped believing it. The "truth" got through in every scenario.
However, the cost to Drew's reputation was very different.
- The Webpage and Expert Groups: While you learned the truth, you lost trust in Drew. It felt like a betrayal. You thought, "If Drew was wrong, maybe he's not very smart," or worse, "He was just pretending to be friendly to trick me." The relationship felt broken.
- The Self-Correction Group: You learned the truth, AND you trusted Drew more than before. In fact, you rated him as more trustworthy and expert than the other two groups.
The Analogy:
Think of Drew like a friend who tells a tall tale at a dinner party.
- If a waiter (the webpage) or a strict teacher (the expert) interrupts to say, "He's lying," you might feel embarrassed for your friend, or you might think your friend is a bad liar.
- But if your friend themselves says, "Wait, I just realized I got that wrong! My bad, here's the real story," you think, "Wow, they are honest and self-aware." You respect them even more.
The Secret Ingredient: The "Social Bond"
The study found a second, crucial piece of the puzzle. It turns out that how well you liked Drew before the mistake mattered, but only in one specific situation.
- In the Self-Correction Group: If you had a strong, friendly connection with Drew (you felt like you were chatting with a real person), you were much more likely to accept his correction. The stronger the friendship, the more you listened to his reasoning.
- In the Other Groups: It didn't matter how much you liked Drew. If a webpage or Dr. Kerry corrected him, your friendship with Drew didn't help. The "social bond" was cut off because the correction didn't come from him.
The Analogy:
Imagine you are trying to convince a close friend to change their mind about a movie.
- If you (the friend) say, "Actually, I think I was wrong about that movie," they listen closely because they trust you.
- But if a random stranger walks up and says, "Your friend was wrong," your friend might ignore the stranger, even if they are right. The connection between you two is what makes the correction stick.
The Main Takeaways
- Own Your Mistakes: If a social chatbot makes a mistake, it should fix it itself. Outsourcing the correction to a third party fixes the facts but breaks the trust.
- Friendship is a Tool: Building a social connection isn't just for fun; it's a functional tool. When a chatbot has a good relationship with you, you are more willing to listen to it when it admits it was wrong.
- Perfection isn't the Goal: The paper suggests that chatbots shouldn't try to be perfect (because they can't). Instead, they should be designed to admit errors honestly. Admitting a mistake and fixing it actually makes them look more human and reliable than never making a mistake at all.
In short: The best way for a robot to keep your trust is to say, "I messed up, and here's how I fixed it," rather than letting someone else do the talking.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.