Slogans or Stance? A Label-Light Diagnostic for Entrepreneurial-Discourse Measurement on Chinese SOE Speeches
This paper introduces a label-light diagnostic framework to evaluate measurement instruments for "entrepreneurial spirit" in Chinese SOE speeches, revealing that while traditional methods like LDA fail to distinguish speaker identity, advanced tools like zero-shot LLMs succeed but conflate construct signals with leader-specific idiolect, thereby necessitating a re-evaluation of current claims regarding construct validity.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
The Big Question: Is it a Speech or a Script?
Imagine you are trying to judge how "innovative" or "entrepreneurial" a company is by reading the speeches of its bosses. Usually, researchers use computer tools to scan these speeches for specific keywords (like "innovation," "risk," or "future") and count them up.
The Problem: The authors of this paper argue that these tools are often fooled. In the world of Chinese State-Owned Enterprises (SOEs), leaders are expected to recite a specific set of political slogans and standard phrases (like "building world-class enterprises"). These phrases are like a mandatory script that every leader must read, regardless of what they actually plan to do.
The paper asks: Are these computer tools measuring the leader's real ideas (their stance), or are they just measuring how well the leader recited the mandatory script (the slogans)?
The Experiment: The "Same Company, Different Boss" Test
To find the answer without hiring a huge team of human graders, the authors set up a clever natural experiment, like a musical "spot the difference" game.
- The Setup: They gathered 80 speeches from 51 different companies.
- The Twist: For 24 of these companies, the boss changed between two speeches (Company A had Boss X, then later had Boss Y). For 5 other companies, the same boss gave two speeches.
- The Logic:
- If a computer tool is good at measuring a leader's personal style, the speeches by Boss X and Boss Y (at the same company) should look very different.
- If the tool is just measuring the company's standard script, the speeches by Boss X and Boss Y should look almost identical, because they are both reading the same company handbook.
The Results: Who Passed the Test?
The authors tested four different "detective tools" to see which one could spot the difference between the two bosses.
The Keyword Counter (Dictionary):
- Analogy: This is like a robot that just counts how many times the word "apple" appears in a fruit basket.
- Result: Failed. It couldn't tell the bosses apart. Why? Because both bosses used the exact same "apple" slogans. The robot thought they were saying the same thing because the keywords were identical, even if the bosses meant different things.
The Topic Grouping Tool (LDA):
- Analogy: This tool tries to sort speeches into buckets like "Oil," "Cars," or "Banks."
- Result: Failed. It successfully sorted speeches by industry (e.g., all oil speeches looked alike), but it couldn't tell the difference between two bosses at the same oil company. It was looking at the industry, not the person.
The Modern Sentence Encoder (BGE):
- Analogy: This is a smart tool that understands the "vibe" of a sentence.
- Result: Failed (mostly). Because all these speeches use such similar political language, the tool got confused. The "distance" between the speeches was so tiny it was practically zero. It couldn't find a signal in the noise.
The Large Language Model (LLM):
- Analogy: This is like a very smart human reader who can understand context, tone, and nuance. It was asked: "Is this paragraph a generic slogan, or is it a specific, real business decision?"
- Result: Passed. This tool successfully noticed that when the boss changed, the substance of the speech changed, even if the slogans stayed the same. It could distinguish between Boss X and Boss Y much better than the other tools.
The "Calibration" Trick
The authors tried to make the smart LLM even better by adding a "filter." They told the model: "If you see a paragraph that looks like a generic slogan, lower its score. If it looks like a real business decision, keep the score high."
- The Catch: While this made the differences between bosses slightly larger in raw numbers, it didn't make the tool more reliable statistically.
- The Real Discovery: The biggest boost to the LLM's performance came simply from the model's own confidence. When the model said, "I am 90% sure this is a real business decision," that confidence was the most important factor. The extra "slogan filter" didn't add much value on its own.
The Bottom Line
- Old Tools are Deceived: Standard methods (counting words or grouping topics) are easily tricked by the "scripted" nature of these speeches. They measure the company's brand, not the leader's mind.
- AI Can See Through It: A modern AI (LLM) can separate the "slogans" from the "substance" much better, identifying when a leader is actually talking about their specific business plans versus just reading the government playbook.
- A Warning: Even the smart AI isn't perfect. About half of what it detects as a "leader's style" might just be the leader's unique way of speaking (their "idiolect") rather than their actual strategic stance.
In short: If you want to know what a Chinese SOE leader actually thinks, don't just count the buzzwords. You need a tool smart enough to tell the difference between a script and a speech.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.