Quantifying Omitted Variable Bias in Nonlinear Instrumental Variable Estimators
This paper develops a framework using double machine learning to quantify omitted variable bias in nonlinear instrumental variable estimators, demonstrating through a U.S. job training experiment that while first-stage compliance estimates are robust, treatment effects are more sensitive to unobserved confounding, particularly for males.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
Imagine you are a detective trying to solve a mystery: Did a specific job training program actually help people earn more money?
You have a clue: the government randomly offered the program to some people (the "Instrumental Variable"). You want to compare the earnings of those who took the program against those who didn't. But here's the problem: people who chose to take the program might have been more motivated, smarter, or had better connections than those who didn't. These hidden traits are the "Omitted Variables"—the invisible factors that mess up your math and make it look like the program worked (or didn't work) when it might not have.
This paper, written by Yu-Min Yen, is like a new super-sensor for your detective work. It teaches you how to measure exactly how much your "hidden clues" (the omitted variables) could be messing up your conclusion, especially when the math gets complicated (nonlinear).
Here is the breakdown of the paper's ideas using simple analogies:
1. The Problem: The "Short" vs. "Long" Recipe
Imagine you are trying to bake a perfect cake (the true effect of the program).
- The "Long" Recipe: You have every single ingredient, including the secret spice (the hidden variable ) that makes the cake taste amazing. If you use this, you get the true answer.
- The "Short" Recipe: You only have the ingredients you can see (the data you have, ). You are missing the secret spice.
Usually, economists just bake the cake with the "Short" recipe and hope for the best. But this paper asks: "How much does the missing spice change the taste?"
The paper looks at three different ways to bake this cake (three types of statistical models):
- PLIVM: A standard linear regression (like a basic cake).
- LATE: Looking only at the people who actually followed the rules (the "Compliers").
- LATT: Looking only at the people who actually took the program (the "Treated").
The paper says: "We can't see the secret spice, but we can calculate a Range of Possibilities for how much the cake might taste different."
2. The Solution: The "Sensitivity Dial"
Since we don't know the secret spice, the authors introduce a Sensitivity Dial.
Imagine a dial on your oven. You can turn it to say, "Okay, let's assume the missing spice is as important as the flour," or "Let's assume it's as important as the sugar."
- If you turn the dial up (assuming the missing variable is very powerful), the range of possible answers gets wider.
- If you turn it down (assuming the missing variable is weak), the range gets tighter.
The paper provides a mathematical framework to turn this dial and see how the "Confidence Interval" (the range of likely answers) stretches or shrinks. It essentially says: "Even if we are wrong about the hidden variables, here is the worst-case scenario and the best-case scenario."
3. The Tool: "Double Machine Learning"
To do this math, the authors use a high-tech tool called Double Machine Learning (DML).
- Analogy: Imagine you are trying to predict the weather. You have a million variables (humidity, wind speed, barometric pressure, the color of the sky, the number of birds flying). A human can't process all that.
- The Machine: You use a super-smart computer (Machine Learning) to figure out which variables matter most.
- The "Double" Part: The computer does this twice. First, it predicts the outcome based on what we can see. Second, it predicts the outcome based on what we can't see (simulated). By comparing the two, it isolates the "bias" caused by the missing data.
This allows the researchers to handle huge amounts of data (like age, race, education, marital status) without getting confused, ensuring their "Sensitivity Dial" is accurate.
4. The Real-World Test: The Job Training Experiment
The authors tested their new method on real data from the U.S. Job Training Partnership Act (JTPA). They looked at whether the program helped men and women earn more money.
The Results (The Plot Twist):
- The "Compliance" (First Stage): The estimate of how many people actually joined the program when offered was very robust. Even if there were hidden variables, the math held up. It's like saying, "We are 100% sure that the offer actually got people to show up."
- The "Effect" (Treatment):
- For Women: The program definitely worked. Even after accounting for the worst-case hidden variables, women still earned significantly more. The "cake" tasted better, no matter what secret spice was missing.
- For Men: The results were fragile. When they turned the "Sensitivity Dial" to account for hidden variables, the statistical significance disappeared. It's like saying, "It looked like the program helped men, but if there was one hidden factor we missed (like a specific personality trait), that result might be a fluke."
5. Why This Matters
Before this paper, if you wanted to know if your results were "real" or just a fluke caused by missing data, you had to guess. You might say, "I think the hidden variable isn't that important," but you had no way to prove it.
This paper gives you a calculator.
- It tells you: "If the hidden variable is as strong as 'Education,' your result is still valid."
- It tells you: "If the hidden variable is as strong as 'Motivation,' your result falls apart."
Summary
This paper is a stress test for economic conclusions. It admits that we never have perfect data (we always miss some variables). Instead of ignoring this, it builds a framework to quantify exactly how much that missing data could hurt your conclusion.
In the case of the job training program, it revealed a crucial truth: The program was a clear success for women, but the evidence for men is shaky and depends entirely on what we didn't measure. This helps policymakers make smarter decisions by knowing which results are solid and which are fragile.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.