← Latest papers
🧬 biology

Simulating is not always understanding: When model complexity obscures biology

The paper argues that true biological understanding stems not from increasing model complexity, but from maintaining a favorable ratio of experimental constraints to free parameters and prioritizing systematic, interpretable analyses that reveal how cellular behaviors emerge from underlying mechanisms.

Original authors: Lendert Gelens, Alejandro Fábregas-Tejeda, Grant Ramsey, Sylvia Wenmackers, Bart Smeets

Published 2026-08-10
📖 9 min read🧠 Deep dive

Original authors: Lendert Gelens, Alejandro Fábregas-Tejeda, Grant Ramsey, Sylvia Wenmackers, Bart Smeets

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). ⚕️ This is an AI-generated explanation of a preprint that has not been peer-reviewed. It is not medical advice. Do not make health decisions based on this content. Read full disclaimer

Imagine you are trying to understand how a car works. You could take the engine apart, lay out every single bolt, piston, and spark plug on a table, and then put it all back together. If the car starts running, you've successfully simulated the engine. But did you actually understand why it runs? Maybe you just got lucky with the order you put the parts back in. Or maybe you could have built a working engine with half as many parts, but you didn't know it because you were too busy counting bolts. This is the dilemma facing modern cell biology. Scientists are building incredibly detailed computer models of cells—digital twins that track thousands of molecules, genes, and chemical reactions. These models are like massive, high-definition video games of life. But there is a growing worry: just because a computer simulation can mimic a cell's behavior perfectly, does that mean we truly understand the rules of the game? The paper you are about to read tackles this question, arguing that sometimes, adding more detail actually makes the mystery harder to solve, not easier.

The authors of this paper, a team of biologists and philosophers of science, are sounding an alarm about "model complexity." They argue that in the rush to build the most detailed, "whole-cell" simulations possible, scientists might be losing the ability to explain why things happen. Think of a model as a recipe. A simple recipe with three ingredients might tell you exactly why a cake rises (the baking soda reacting with the vinegar). But if you write a recipe with 5,000 ingredients, including "a pinch of stardust" and "three drops of morning dew," and the cake still rises, you haven't learned anything new. You can't tell which of the 5,000 ingredients actually mattered. The paper suggests that true understanding doesn't come from piling on more data; it comes from having a model where the "knobs" (the numbers you can adjust) are tightly controlled by real-world experiments, so you can see exactly which knobs make the machine tick.

The Trap of the "Perfect" Simulation

The paper starts by pointing out a common trap: Simulation is not the same as Understanding. Imagine you are trying to guess the combination to a safe. If you have a robot that tries every single combination from 0000 to 9999, it will eventually open the safe. The robot has "simulated" the opening of the safe. But the robot doesn't know the combination; it just knows that one of those numbers worked. In cell biology, a complex model might "fit" the data perfectly, meaning it reproduces what scientists see in the lab. But if the model has too many free numbers (parameters) that can be tweaked to make it work, it's like the robot trying every combination. It proves the model can work, but it doesn't prove the model is the right explanation.

The authors argue that for a model to teach us something, it needs to do more than just copy what we already know. It needs to:

  1. Make a new prediction that we haven't seen yet.
  2. Reveal a surprising connection between two things we thought were unrelated.
  3. Fail in a specific way when we change a part, showing us that part was essential.

If a model is so complex that it can be tweaked to fit almost anything, it fails at all three of these. It becomes a "black box" that gives the right answer for the wrong reasons.

The "Sloppiness" Problem

Here is where the paper gets really interesting with a concept called parameter sloppiness. Imagine you are tuning a radio. If you have a simple radio with just one knob for volume, it's easy to find the right setting. But imagine a radio with 1,000 knobs, and you can turn them all to different settings to get the same clear song. This is "sloppiness." In these complex biological models, scientists often find that they can change dozens of numbers by huge amounts, and the model still produces the exact same result.

The paper explains that this is actually a double-edged sword.

  • The Good News: The model is great at predicting. Because the result doesn't change much when you tweak the numbers, the model is robust. It will likely give you the right answer even if you don't know the exact value of every single part.
  • The Bad News: The model is terrible at explaining. If you can change 100 numbers and get the same result, you have no idea which of those 100 numbers is actually doing the work. You can't say, "Ah, this specific protein is the cause!" because the model says, "Actually, it could be any of these 50 proteins, or a mix of them."

The authors use a vivid example of a cell cycle model (the process a cell goes through to divide). They compared a "detailed" model with 13 adjustable numbers to a "minimal" model with only 3. Both models could make the cell divide in the right amount of time. However, when they tested how much the numbers could wiggle around:

  • The detailed model was "sloppy." It could work with about 85% of all possible number combinations. It was impossible to tell which numbers were the "real" ones.
  • The minimal model was "tight." It only worked with about 2% of the combinations. Because it was so constrained, the scientists could actually learn something about the system.

The paper argues that adding more biological detail (more proteins, more reactions) without adding more experimental data to pin those numbers down just makes the model "sloppier." It's like trying to solve a puzzle with a million pieces, but you only have the picture on the box to guide you. You might fit the pieces together, but you won't know if you did it the right way.

The Solution: Build a Ladder, Not a Wall

So, what should scientists do? The paper suggests we stop trying to build the biggest, most complex model possible as the final goal. Instead, we should treat complex models as a starting point for a journey of discovery.

The authors propose two main strategies to turn a "simulation" into "understanding":

  1. Build Model Hierarchies (The Ladder): Instead of starting with a massive model, scientists should build a ladder of models. Start with the simplest version that can do the job. Then, add complexity step-by-step. At each step, ask: "Did we need to add this new part to make it work?" If the simple model works just as well, the complex part wasn't necessary for that specific behavior. This helps strip away the "irrelevant" details and find the core rules. The paper points out that in some areas, like modeling how cells stick together in tissues, scientists have already done this. They use simple models with just a few rules to explain complex behaviors like how tissues flow or jam, and because the rules are simple, they can explore every possibility and truly understand the mechanics.

  2. Do the "Dynamical Analysis" (The Map): Scientists need to stop just running the simulation once and seeing if it looks right. They need to map out the "phase diagram." Imagine a map that shows all the different states a system can be in. By systematically changing the numbers in the model and seeing how the behavior changes, scientists can find the "tipping points" (bifurcations). This tells them what controls the system and what happens if you push it too hard. The paper notes that while this is common in simple models, it is rarely done in massive "whole-cell" models because they are too computationally expensive to run thousands of times. But without this map, the model is just a demonstration, not an explanation.

The Machine Learning Twist

The paper also touches on the rise of Machine Learning (AI) in biology. AI models are getting really good at predicting what a cell will do based on huge amounts of data. But the authors warn that these AI models face the same problem. They are often "sophisticated interpolators"—they are really good at guessing the answer based on patterns in the data, but they don't necessarily know why the answer is what it is.

The authors suggest that if AI can build models quickly and cheaply, the real challenge shifts from building the model to analyzing it. Just because a computer can generate a model that fits the data doesn't mean the model is useful. We still need to do the hard work of checking if the model's "knobs" are constrained and if we can explain the cause-and-effect relationships.

The Bottom Line

The paper concludes with a call to change the "ethos" of biological modeling. We shouldn't aim to include every single molecule in a cell just to say we did it. Instead, we should aim for understanding.

  • Complexity isn't the goal: A model with millions of parts isn't automatically better than one with a few.
  • Constraints are key: A model is only useful if the numbers inside it are pinned down by real experiments.
  • Simplicity reveals truth: Sometimes, the best way to understand a complex system is to find the simplest version that still works, and then see what happens when you add things back in.

The authors aren't saying big models are useless. They are saying that building a "whole-cell" simulation is just the first step, like gathering all the ingredients for a cake. The real work—the part that gives us understanding—happens when we bake the cake, taste it, and figure out exactly which ingredient made it rise. Until we do that analysis, we might just be simulating the cell without really knowing how it works.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →