← Latest papers
🧬 biology

Querying Counterfactuals on Tissue Graphs with Supervised Disentanglement

This paper introduces Cellina, a supervised disentanglement framework that formalizes tissue graph counterfactuals by separating a cell's intrinsic state from its spatial context to accurately predict expression changes under rewired or modified neighbor interactions, outperforming existing methods on large-scale spatial transcriptomics datasets.

Original authors: Abdul Moeed, Stefan Schrod, Martin Rohbeck, Marc Jan Bonder, Pavlo Lutsik, Oliver Stegle, Daniel Dimitrov

Published 2026-06-09
📖 5 min read🧠 Deep dive

Original authors: Abdul Moeed, Stefan Schrod, Martin Rohbeck, Marc Jan Bonder, Pavlo Lutsik, Oliver Stegle, Daniel Dimitrov

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). ⚕️ This is an AI-generated explanation of a preprint that has not been peer-reviewed. It is not medical advice. Do not make health decisions based on this content. Read full disclaimer

The Big Picture: The "Cell Neighborhood" Problem

Imagine a city where every person (a cell) has a unique personality and a specific job. In biology, we want to predict how a person will act if they move to a different neighborhood.

  • The Old Way: Previous computer models treated people like they were living in isolation. They assumed that if you gave everyone the same "stimulus" (like a new law or a virus), everyone would react the same way, regardless of who their neighbors were.
  • The Reality: In a living body, your behavior is heavily influenced by your neighbors. A person living next to a bakery might smell bread and feel hungry; a person next to a construction site might feel stressed. In biology, a cell's "mood" (gene expression) is shaped by the cells right next to it.

The paper argues that to predict how a cell will change, we can't just look at the cell itself; we have to simulate what happens if we swap its neighbors or change what its neighbors are saying.

The Solution: "Cellina" (The Neighborhood Simulator)

The authors built a tool called Cellina. Think of Cellina as a super-smart translator that can separate a person's "inner self" from their "environment."

  1. The Two Parts of a Person:

    • The Inner Self (Intrinsic): Who you are at your core (e.g., "I am a baker"). This doesn't change just because you move.
    • The Environment (Extrinsic): How your surroundings affect you (e.g., "I smell like bread because I live next to a bakery").
  2. The Magic Trick (Disentanglement):
    Many AI models mix these two things up. If you move a baker to a construction site, a bad model might think the baker became a construction worker. Cellina is special because it uses "supervised disentanglement."

    • Analogy: Imagine you have a photo of a person. Cellina uses a special filter to separate the person from the background. It learns to recognize the person even if the background changes, and it learns to recognize the background even if the person changes.
    • How it learns: It uses a "teacher" (cell type labels) to say, "This part of the data is the person's identity," and a "trickster" (adversarial training) to say, "Make sure the 'person' part doesn't accidentally learn anything about the 'neighborhood'."

The Two Types of Experiments

The paper tests Cellina with two specific "What if?" scenarios, which they call Counterfactuals:

1. Edge Perturbation (The "Moving House" Test)

  • The Question: "What would this cell look like if we physically moved it to a different neighborhood and gave it a completely new set of neighbors?"
  • The Analogy: Imagine taking a baker and moving them to a construction site, replacing their old neighbors with construction workers. Cellina predicts how the baker's behavior changes just by being surrounded by these new people.
  • The Result: Cellina was much better at this than other tools. It successfully predicted that the baker would start acting differently (e.g., smelling like sawdust) without losing their identity as a baker.

2. Node Perturbation (The "Changing the Message" Test)

  • The Question: "What if the neighbors stay in the same house, but they start shouting different things?"
  • The Analogy: The baker stays next to the construction workers, but instead of shouting about drills, they start shouting about blueprints. Cellina predicts how the baker reacts to this change in conversation without moving.
  • The Result: Cellina could also simulate this, predicting how specific biological pathways (like a specific gene program) would turn on or off based on these new signals.

Why This Matters (According to the Paper)

The authors tested Cellina on two massive datasets:

  1. Colorectal Cancer: Looking at cells in healthy colon tissue vs. cancerous tissue.
  2. Mouse Brain: Looking at different regions of a mouse brain.

The Findings:

  • Better Predictions: Cellina beat all other existing methods (like scGen, CPA, and MintFlow) in predicting how cells would change in these "virtual experiments."
  • Finding Hidden Patterns: Because Cellina separates the "neighborhood" from the "cell," it found hidden sub-groups within the cancer tissue that other methods missed. It realized that not all cancer cells are in the same "environment," even if they are in the same tumor.
  • Scalability: It can handle millions of cells without crashing, which is a big deal for modern biology.

What Cellina Does Not Do (Based strictly on the text)

  • It is not a crystal ball for cures: The paper explicitly states these are "in-silico" (computer simulation) predictions. They are "generative hypotheses" that need to be tested in a real lab (wet-lab validation) before we know if they are true.
  • It doesn't replace doctors: The authors do not claim this tool can diagnose patients or prescribe medicine yet. It is a research tool to help scientists understand how cells interact.
  • It relies on labels: To work its magic, Cellina needs to be told what "type" of cell it is looking at (e.g., "This is a T-cell") and what "region" it is in (e.g., "This is the tumor"). It learns from these labels to separate the identity from the environment.

Summary

Cellina is a new computer program that treats tissue like a graph of connected neighbors. It learns to separate a cell's identity from its surroundings. By doing this, it can answer "What if?" questions: "What if this cell had different neighbors?" or "What if its neighbors said something else?" It does this better than any previous tool, helping scientists understand the complex social life of cells in tissues like tumors and brains.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →