← Latest papers
📊 statistics

Sensitivity analysis for causal mediation: bridge score, sharp sensitivity bounds, and calibration

This paper introduces the "bridge score" to derive sharp sensitivity bounds for unobserved mediator-outcome confounding in causal mediation analysis and provides calibration methods and a Bayesian algorithm to operationalize these bounds for robust inference.

Original authors: Yuki Ohnishi, Fan Li

Published 2026-05-19
📖 5 min read🧠 Deep dive

Original authors: Yuki Ohnishi, Fan Li

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer

Imagine you are trying to figure out why a specific treatment (like a new medicine or a political message) works. You suspect it works through a specific "middleman" (a mediator). For example, a negative news story might make people feel angry (the mediator), which then makes them support stricter immigration laws (the outcome).

To prove this, scientists usually assume that once they account for the anger, there are no other hidden factors secretly linking the news story to the policy support. But what if there are hidden factors? Maybe the people who got the angry news also happened to be naturally more anxious, and that anxiety drives both the anger and the policy support. If you don't account for this, your results are wrong.

This paper introduces a new, more robust way to test how much those hidden factors could be messing up your results. Here is the breakdown using simple analogies:

1. The Problem: The "Hidden Puppeteer"

In standard analysis, researchers pretend the "middleman" (the mediator) is innocent. But in reality, there might be a "hidden puppeteer" (an unmeasured confounder) pulling strings on both the mediator and the outcome.

  • The old way: Researchers used to guess a single number (like a correlation coefficient) to represent this hidden puppeteer. The problem is that this guess was often too simple or too rigid, like trying to describe a complex storm with a single wind speed number.
  • The new way: This paper says, "Let's not guess a single number. Let's build a better map to see exactly how strong that puppeteer could be."

2. The Solution: The "Bridge Score"

The authors invent a new tool called the Bridge Score.

  • The Analogy: Imagine you are trying to cross a river. On one side is the "Control Group" (people who didn't get the treatment), and on the other is the "Treatment Group" (people who did).
  • Usually, you look at the people on both sides separately. But the Bridge Score looks at a specific person and asks: "If this person had been in the Control group, how likely were they to have this specific mediator value? And if they were in the Treatment group, how likely were they to have it?"
  • It creates a two-dimensional "ID card" for every person based on these two probabilities. This ID card acts as a "bridge" that balances the two groups perfectly.
  • Why it helps: Instead of trying to account for every single detail about a person (age, income, location, etc.) to find the hidden puppeteer, the Bridge Score condenses all that information into this single, powerful ID card. It's like using a high-resolution fingerprint instead of a blurry photo to identify someone.

3. The "Sharp Bound": The Safety Net

Once they have this Bridge Score, the authors calculate a "Sharp Bound."

  • The Analogy: Think of this as a safety net with a very precise ceiling. If you assume the hidden puppeteer is this strong, the error in your results can't possibly be higher than this ceiling.
  • Why "Sharp"? Many old methods gave a safety net that was so wide (like a net covering the whole sky) that it wasn't useful. This new net is tight. It tells you the exact maximum amount of error possible given your assumptions. It's the difference between saying "The error could be anywhere between 0 and 100%" versus "The error is definitely between 0 and 12%."

4. Calibration: "The Ruler Test"

The biggest challenge is: "How do we know how strong the hidden puppeteer actually is?" We can't measure it directly.
The paper offers two ways to calibrate (measure) this strength using things we can see:

  • Method A: The "Benchmark" (The Known Ruler)

    • Imagine you have a known variable, like "Income." You know how much Income affects the outcome.
    • You ask: "Is our hidden puppeteer stronger than Income? Is it twice as strong?"
    • The Innovation: The authors suggest using the Rank of the income (e.g., "Is this person in the top 10%?") rather than the raw dollar amount. This is like measuring height in "inches" vs. "centimeters." Using the rank makes the test fair and stable, regardless of whether you measure money in dollars or euros. It prevents the results from changing just because you changed the units.
  • Method B: The "Residual Budget" (The Leftover Energy)

    • Imagine you have a bucket of "unexplained energy" (variation in the outcome that your model couldn't explain).
    • You ask: "How much of this leftover energy could the hidden puppeteer be stealing?"
    • You set a limit: "The puppeteer can't steal more than 20% of the leftover energy." This gives you a concrete, data-driven limit on how much error is possible.

5. The Result: A Clearer Picture

By combining the Bridge Score (the ID card), the Sharp Bound (the tight safety net), and the Calibration (the ruler test), the authors provide a way to say:

"Even if there is a hidden puppeteer, and even if it is as strong as [Income] or [20% of the leftover energy], our conclusion about the treatment effect is still likely to hold true. Here is the exact range of how much it could change."

Summary

This paper doesn't just say "be careful about hidden factors." It builds a precision instrument to measure exactly how much those hidden factors could hurt your study.

  • It uses a Bridge Score to simplify complex data.
  • It creates a Sharp Bound to define the exact limits of error.
  • It uses Calibration to let researchers use real-world data (like income or leftover variation) to set realistic limits on their assumptions.

This allows scientists to be much more confident in their conclusions about how and why treatments work, even when they can't measure every single hidden factor in the world.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →