← Latest papers
💬 NLP

Expect the Unexpected? Testing the Surprisal of Salient Entities

This paper challenges the Uniform Information Density hypothesis by demonstrating that globally salient entities in discourse exhibit significantly higher surprisal than non-salient ones, while simultaneously reducing the surprisal of surrounding content, thereby revealing global entity salience as a key mechanism shaping information distribution across different genres.

Original authors: Jessica Lin, Amir Zeldes

Published 2026-04-14
📖 5 min read🧠 Deep dive

Original authors: Jessica Lin, Amir Zeldes

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer

The Big Idea: The "Main Character" Effect

Imagine you are reading a mystery novel. The story is full of clues, red herrings, and side characters. But there is one Main Character (let's call them "Detective Smith") who is the heart of the story.

The researchers wanted to know two things about Detective Smith:

  1. Is Detective Smith surprising when they show up? (Do we gasp when they enter the room?)
  2. Does Detective Smith make the rest of the story easier to predict? (Once we know Smith is there, do we know exactly what kind of trouble is about to happen?)

This paper investigates how the "importance" of a person or thing in a story (called Global Salience) changes how predictable the rest of the text feels to a computer (or a human brain).


Part 1: The Theory (Uniform Information Density)

Scientists have a theory called Uniform Information Density (UID). Think of it like a river flowing at a steady speed.

  • The Theory: Speakers and writers try to keep the flow of new information smooth and even. They don't want a sudden flood of confusing words followed by a dry spell of boring words. They want a steady stream.
  • The Reality: The river isn't perfectly smooth. Sometimes, grammar rules or the way we speak force a "bump" in the water.

The researchers asked: Does the "Main Character" of a story create a bump in this river?

Part 2: The Experiment (The 70,000 Clues)

The researchers used a massive library of 16 different types of writing (like academic papers, news stories, fiction, and casual conversations). They looked at over 70,000 mentions of people and things.

They defined a "Salient Entity" (a Main Character) as something that would appear in a summary of the text.

  • Analogy: If you asked 5 different people to summarize a movie, and they all mentioned "The Villain," then "The Villain" is a Salient Entity. If only one person mentioned a random background extra, that extra is Non-Salient.

Part 3: The Three Surprising Findings

The researchers ran three experiments to see how these "Main Characters" affect the flow of information.

1. The "Main Character" is actually more surprising when they appear.

  • The Finding: When a "Main Character" (like "The President" in a political speech) is mentioned, it is actually more surprising than a random word.
  • The Analogy: Imagine you are walking down a street. You see a random passerby (non-salient). That's boring; you expect it. But then, suddenly, the Mayor walks by (salient). That is a surprise! You have to pay attention.
  • Why? Because Main Characters carry a lot of weight. When they show up, it's a big deal, so the brain (or computer) has to work harder to process that specific moment.

2. The "Main Character" acts as a lighthouse.

  • The Finding: Even though the Main Character is a surprise when they arrive, everything that comes after them becomes easier to predict.
  • The Analogy: Think of the Main Character as a lighthouse.
    • When the lighthouse beam hits you, it's bright and startling (high surprisal).
    • But once you see the lighthouse, you know exactly where the rocks are and where the safe path is. The rest of the journey becomes predictable.
    • If you mention "The President," you immediately expect words like "speech," "policy," or "economy." If you mention a random name like "Bob," the next words could be anything.
  • The Result: Salient entities create "troughs" (dips) in the confusion. They make the surrounding text feel very smooth and predictable.

3. The "Genre" matters (The Party vs. The Lecture).

  • The Finding: This "lighthouse effect" works best in organized texts (like textbooks or news) and fails in chaotic texts (like casual chats).
  • The Analogy:
    • The Lecture (Academic/News): Everyone is listening to one topic. If the speaker mentions the "Main Topic," the audience knows exactly where the conversation is going. The lighthouse works perfectly.
    • The Party (Conversation/Vlogs): People are jumping from topic to topic. One minute we are talking about pizza, the next about a cat, then about the weather. Even if you mention a "Main Character" (like "My Mom"), she might be talking about the cat, not the pizza. The lighthouse beam is spinning wildly, and it doesn't help you predict what comes next.

The "Minimal Pair" Trick (How they proved it)

To make sure they weren't just tricking themselves, the researchers used a clever trick called a Minimal Pair.

Imagine they took a sentence: "The prevalence of discrimination across racial groups..."

  • Scenario A: They put the word "Discrimination" (a Main Character) in front of it.
  • Scenario B: They put the word "Psychologists" (a random side character) in front of it.

They asked a computer: "Which sentence is easier to guess?"

  • Result: The computer guessed the sentence much faster after "Discrimination" than after "Psychologists." This proved that the importance of the word, not just its length or grammar, was making the text predictable.

The Conclusion: Why This Matters

This paper changes how we think about how humans write and speak.

  1. We aren't robots: We don't just try to keep information perfectly smooth. We intentionally create "bumps" (surprises) when we introduce important characters.
  2. We are guides: Once we introduce those important characters, we use them to guide the listener, making the rest of the story flow smoothly.
  3. Context is King: This only works if the story stays on topic. In a chaotic conversation, the "Main Character" doesn't help as much because the topic keeps changing.

In short: The most important people in a story are the most surprising when they arrive, but they are also the best guides for where the story is going next.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →