Statistical tests for bivariate spatial association across multi-omics data with disjoint coordinates
This paper introduces the R-package `sbivar`, which provides a suite of modified statistical tests and variance estimators to rigorously assess bivariate spatial associations between multi-omics modalities with disjoint coordinates while properly accounting for spatial autocorrelation and high-dimensional computational challenges.
Original paper licensed under CC BY 4.0 (https://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of a preprint that has not been peer-reviewed. It is not medical advice. Do not make health decisions based on this content. Read full disclaimer
Imagine you are a detective trying to solve a mystery inside a living city. This city is a piece of biological tissue, and the "citizens" are molecules like genes and metabolites. In the past, scientists could only take a blurry photo of the whole city or a tiny, high-definition photo of just one street corner. But now, we have super-powered microscopes that can map out exactly where different types of molecules are living within the tissue. The big question is: Do certain molecules like to hang out together? If a specific gene and a specific chemical are found in the same neighborhood, they might be working together on a secret project, like building a cell wall or fighting an infection. Finding these "buddy pairs" helps us understand how life works, how diseases start, and how to fix them.
However, there's a catch. The different microscopes used to map these molecules often have different resolutions. One might take a picture with tiny, pixel-perfect dots, while the other takes a picture with slightly larger, fuzzier dots. Even if you line up the two pictures perfectly, the dots from one picture rarely land exactly on top of the dots from the other. It's like trying to match a grid of tiny LEGO bricks to a grid of slightly larger Jenga blocks; they just don't line up one-to-one. Furthermore, molecules in biology aren't scattered randomly like marbles on a floor; they tend to clump together in patterns, a phenomenon scientists call "spatial autocorrelation." If you ignore these clumps and the mismatched grids, your math can get very confused, leading you to think two molecules are best friends when they are actually just strangers who happened to be in the same general area.
This is where the new paper steps in with a fresh set of tools. The authors, Stijn Hawinkel and his team, realized that many existing methods for finding these molecular buddy pairs were using the wrong math. They were treating the data as if the molecules were scattered randomly and perfectly aligned, which led to a lot of false alarms—finding "friendships" that didn't actually exist. The team argues that you need a method that can handle mismatched grids (disjoint coordinates) and still respect the natural clumping of molecules.
To solve this, they developed a new software package called sbivar (which you can find on GitHub). They didn't just invent one new trick; they adapted three different statistical approaches to work on these messy, mismatched maps:
- The "Neighborhood Watch" (Bivariate Moran's I): Imagine you are checking if two types of people are hanging out in the same neighborhood. Instead of needing them to stand on the exact same spot, this method looks at how close they are to each other. The authors tweaked this old method to calculate the odds of them being together while accounting for the fact that people naturally clump together. They created a new way to measure the "noise" in the data so they don't get fooled by the clumping.
- The "Smooth Map" (Generalized Additive Models or GAMs): This method is like drawing a smooth, wavy blanket over the data points to see the big picture. It creates a smooth surface for each molecule type based on where they are, even if the dots are on different grids. Then, it compares the two smooth blankets to see if they rise and fall in the same places. The team figured out how to measure the uncertainty of these blankets so they know if the match is real or just a fluke.
- The "Distance Detective" (Bivariate Gaussian Processes): This is a more complex mathematical model that treats the tissue like a continuous field where the relationship between molecules changes based on distance. It's very powerful but can be slow to run on a computer.
The team tested these new tools using computer simulations. They created fake tissue maps with known patterns—some where molecules were truly friends, and some where they were just random neighbors. They found that the old methods often screamed "Friendship!" when there was none, especially when the molecules had natural clumps. In contrast, their new methods (especially the Neighborhood Watch and the Smooth Map) were much better at keeping quiet when there was no friendship and shouting "Friendship!" only when it was real. They also showed that these new tools work well even when the maps are slightly misaligned, which happens often in real experiments.
Finally, they took their new tools to real-world data from human lung and breast cancer samples, and mouse brain samples. They found some exciting new pairs. For instance, in a lung cancer sample, they discovered that a group of genes called Kallikreins were hanging out with a specific fat molecule called PI 34:1. This makes sense biologically because both are involved in a major signaling pathway that cells use to communicate. In the mouse brain, they found a gene (CSRP1) and a fat molecule (palmitoyl-L-carnitine) that were neighbors in the brain's striatum. This pairing hints at how brain cells and their support cells (astrocytes) might be swapping energy sources.
The paper doesn't claim to have solved every mystery in biology, and the authors admit that some of their tools are slower than others. However, they have provided a much more reliable way to find these molecular connections without getting tricked by the messy nature of real-world data. By fixing the math, they hope scientists can now trust their findings about which molecules are truly working together, leading to better insights into how our bodies function and how diseases like cancer might be stopped.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.