What's in a Name? Morphological Shortcuts by LLMs in Pharmacology
This paper reveals that large language models in pharmacology often rely on morphological shortcuts, specifically drug name affixes, to infer properties of fictitious drugs, a behavior localized to early-mid model layers that poses subtle but measurable safety risks due to overgeneralization and unindicated reliance on these cues.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
The Big Idea: The "Suffix" Cheat Code
Imagine you are taking a test on medicine. You don't actually know what a specific drug does, but you notice it ends with the letters "-cillin" (like ampicillin). You remember that penicillin and amoxicillin are antibiotics. So, you guess, "Oh, this must be an antibiotic too!"
You might be right, or you might be wrong. But here is the scary part: Large Language Models (LLMs) are doing this exact same thing, but with 100% confidence.
This paper investigates a "shortcut" these AI models take. Instead of remembering the actual facts about a specific medicine, they look at the shape of the word (its prefixes and suffixes) and guess the answer based on that pattern.
The Experiment: The "Fake Drug" Test
The researchers wanted to see if AI models could be tricked by word shapes. They created three types of words to test on the AI:
- Real Drugs: Actual medicines like Ampicillin.
- Fake Drugs (The Trap): Made-up names that sound real because they use a real medical suffix. For example, they took a nonsense word like "dimi" and added the real antibiotic suffix "-cillin" to make "Dimicillin."
- Nonsense Words: Made-up words with made-up parts, like "Dimiglimto," which have no medical meaning.
The Results:
When the AI was asked about Dimicillin, it didn't say, "I don't know, that's fake." Instead, it confidently said, "Yes, Dimicillin is an antibiotic used to treat bacterial infections."
It was so confident that it treated a completely made-up word exactly the same as a real one, simply because the word ended with the "magic letters" that usually mean "antibiotic."
The Three Key Findings
The paper breaks down exactly how and why this happens:
1. The "Morphological Shortcut" (The Behavior)
The AI isn't "thinking" about the drug; it's pattern-matching.
- Analogy: Imagine a chef who has never cooked a specific dish but knows that any dish ending in "-pizza" usually has cheese and tomato. If you hand them a plate labeled "Dog-pizza," they will confidently describe the cheese and tomato, even though "Dog-pizza" doesn't exist.
- The Risk: In medicine, this is dangerous. If a doctor asks an AI about a fake drug, the AI might invent a fake cure or side effect, and because it sounds so confident, the human might believe it.
2. The "Hidden Reasoning" (The Diagnostic)
The researchers built a tool to see where the AI gets its information. They found that for many drugs, the AI relies mostly on the suffix (the ending) rather than the whole name.
- The Surprise: The AI rarely admits this. When asked why it thinks a drug is an antibiotic, it won't say, "Because it ends in -cillin." It will just give a confident answer as if it knows the facts. It's like a student who guesses the answer because the question looks familiar, but then writes a long essay pretending they studied the topic.
- The Confusion: Because the AI relies on the suffix, it often mixes up real drugs. If two real drugs both end in "-cillin," the AI might accidentally give the side effects of Drug A to Drug B, just because they look alike.
3. The "Brain Scan" (The Mechanism)
The researchers looked inside the AI's "brain" (its internal layers) to see where this shortcut happens.
- The Location: They found that the AI makes this decision very early in its processing, specifically in the early-to-middle layers of its neural network.
- The Mechanism: It's like a security guard at a club who only checks the color of your shirt (the suffix) and lets you in without checking your ID (the specific drug facts). Once the AI sees the "right color shirt," it immediately decides, "This is an antibiotic," and stops looking for more details.
- The Good News: Because they found exactly where this happens in the AI's brain, they showed that it is possible to "patch" or fix this behavior by tweaking those specific layers.
Why This Matters (According to the Paper)
The paper concludes that this isn't just a funny quirk; it's a safety risk.
- It's subtle: The AI sounds very smart and confident.
- It's measurable: We can prove the AI is cheating by looking at its word patterns.
- It's fixable: Since we know where the "shortcut" lives in the AI's code, we can potentially build safety checks to stop it from making up medical facts based on word endings.
In short: The paper warns us that AI models are sometimes "word-surfers" rather than "fact-knowers." In high-stakes fields like medicine, surfing on the shape of a word can lead to confident but completely made-up medical advice.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.