Imperative Interference: Social Register Shapes Instruction Topology in Large Language Models
This paper demonstrates that the imperative mood in system prompts creates language-dependent instruction topologies—cooperative in English but competitive in Spanish—due to varying social register conventions, whereas rewriting instructions in the declarative mood significantly reduces this cross-linguistic variance.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
The Big Idea: Instructions Are Social, Not Just Technical
Imagine you are the captain of a ship. You have a crew (the AI) and a rulebook (the System Prompt).
Usually, we think of a rulebook as a list of technical facts: "The engine is on," "The sails are up." But this paper argues that for AI, a rulebook is actually a social conversation. It's about who is talking to whom, and how they are saying it.
The researchers discovered something weird: The same set of rules works perfectly in English, but falls apart when translated into Spanish.
It's not that the translation is bad. It's that the way the rules are written creates a different social vibe in each language.
The Analogy: The "Commanding Officer" vs. The "Fact Sheet"
To understand why this happens, imagine two different ways to tell a group of people what to do:
The Imperative Style (The Commanding Officer):
- "NEVER touch the red button!"
- "ALWAYS wear your helmet!"
- "DO NOT run!"
- The Vibe: This is loud, authoritative, and direct. In English, if you stack five of these commands, the crew feels a single, strong voice of authority. They all agree: "Okay, the Captain is speaking. We listen."
The Declarative Style (The Fact Sheet):
- "Red button: Do not touch."
- "Helmet status: Required."
- "Running: Prohibited."
- The Vibe: This is calm and factual. It's not a command; it's just a description of how the world works.
The Problem: The "Social Noise" in Spanish
The researchers found that in English, the "Commanding Officer" style works great. The AI listens to all the commands and follows them together.
But in Spanish, the "Commanding Officer" style creates social noise.
- In English: Stacking commands feels like one strong leader giving a clear list of orders. The AI thinks, "Got it, Captain."
- In Spanish: Stacking those same commands feels like five different people shouting orders at once. Because of cultural and linguistic nuances, the "obligatory force" (how much you have to listen) feels different. The AI gets confused. It thinks, "Wait, is Person A telling me to do this, while Person B is telling me not to? Who is the boss?"
Instead of working together, the instructions start fighting each other. Removing one command actually helps the AI perform better because it stops the shouting match.
The Experiment: The "Rewrite" Fix
The researchers tested this with a real-world AI system (Claude Code). They took a prompt with 56 instructions and ran experiments:
- The Baseline: They saw that in English, the instructions worked like a cooperative team. In Spanish, they worked like a competitive brawl.
- The Fix: They took the "Commanding Officer" instructions in Spanish and rewrote them as "Fact Sheets" (Declarative).
- Before: "You must never use this tool!" (Command)
- After: "This tool is disabled." (Fact)
- The Result:
- The Magic: When they rewrote just three of the eleven "Commanding" instructions into "Facts," the entire Spanish system suddenly stopped fighting. The instructions became cooperative again.
- The Spillover: Even the instructions they didn't rewrite started working better. It's like turning down the volume on three shouting people in a room; suddenly, everyone else in the room can hear each other better, too.
Why Does This Matter?
This changes how we think about AI safety and training.
- The "Constitutional AI" Problem: Many AI models are trained using "Constitutional AI," where humans teach the AI rules like "Be helpful," "Be harmless," and "Be honest." These are written as commands (Imperatives).
- The Prediction: If the AI learns these rules in English (where commands work well) but then tries to use them in Spanish (where commands cause confusion), the AI might be "good" in English but "weird" or "unreliable" in Spanish. It's not because the AI doesn't know the language; it's because the social tone of the rules doesn't translate.
The Takeaway
If you want an AI to follow rules perfectly in every language, stop shouting commands.
Instead of saying, "NEVER do X! ALWAYS do Y!", say, "X is disabled. Y is required."
By treating instructions as facts rather than orders, you bypass the cultural confusion. You stop the AI from trying to figure out "who is the boss" and let it just focus on "what is the rule."
In short: The paper proves that AI isn't just a calculator; it's a social being that reacts differently to the tone of voice depending on the language. To get the best results, speak in facts, not commands.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.