BLOSUM Is All You Learn - Generative Antibody Models Reflect Evolutionary Priors
This study demonstrates that the correlation between generative model log-likelihoods and antibody binding affinity arises because these models implicitly learn evolutionary priors, performing on par with simple BLOSUM similarity scores calculated against a known parental binder.
Original paper licensed under CC BY 4.0 (https://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of a preprint that has not been peer-reviewed. It is not medical advice. Do not make health decisions based on this content. Read full disclaimer
Imagine you are trying to design a new key (an antibody) that fits perfectly into a specific lock (a virus or disease target). In the past, scientists used complex, high-tech "AI key-makers" (generative models) to invent these keys. Recently, they noticed something strange: the AI's internal confidence score (its log-likelihood) seemed to predict how well the new key would actually work.
This paper asks a simple question: Why does the AI's confidence score predict success?
Here is the breakdown of their discovery, using everyday analogies:
1. The "Old-School" vs. The "High-Tech"
The researchers compared the fancy AI models to a much simpler, old-school tool called BLOSUM. Think of BLOSUM as a "family resemblance calculator." It looks at a new key design and compares it to the original, proven key (the "parental" sequence) to see how similar they are.
- The Finding: Surprisingly, this simple "family resemblance" score worked just as well as the complex AI at predicting if a new key would fit the lock.
2. The AI is Just Copying the "Family Tree"
The study discovered that the fancy AI models aren't actually learning some secret, magical physics of how antibodies bind. Instead, they are implicitly learning the rules of evolution that are already written into the BLOSUM calculator.
- The Analogy: It's like a student taking a test. The AI isn't inventing new answers; it's just memorizing the patterns of the textbook (evolutionary history) so well that its "confidence score" is basically just a fancy way of saying, "This looks a lot like the things that have worked before."
3. The "Distance" Rule
Both the AI score and the BLOSUM score act like a distance meter.
- If you design a key that is very close to the original, proven key, the score is high, and it's likely to work.
- As you move further away from the original design (making it more different), the score drops, and the chance of it working decreases.
- Essentially, the further you stray from what evolution has already proven works, the less likely you are to succeed.
4. What Doesn't Work (The "Generic" Approach)
The researchers tried using other tools that don't look at the specific "parent" key, but instead look at a generic list of average keys (using things called PWMs and PSSMs).
- The Result: These generic tools were inconsistent. Sometimes they worked, sometimes they didn't. It's like trying to guess the right key by looking at a pile of random keys from a junk drawer rather than comparing your new design to the specific lock you are trying to open.
5. The "Consensus" Trap
Finally, they tried replacing the specific "parent" key with a "consensus" key (an average of all keys ever made).
- The Result: The correlation vanished. This proves that the success of these scores depends entirely on knowing the specific context of the original, working key. You can't just use a generic average; you need to know exactly what you are building upon.
The Bottom Line
The paper concludes that these complex AI models are largely acting as sophisticated mirrors of evolutionary history. The reason their scores predict success is that they are measuring how close a new design is to a known, working ancestor. Therefore, simple, interpretable tools that measure this "family resemblance" can be just as effective as the complex AI for ranking antibody designs, offering a clearer window into why a design might work.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.