← Latest papers
🧬 biology

RECIPE bridges transcriptomics and proteomics with deep graph learning on Ribo-seq data

The paper introduces RECIPE, a graph neural network framework that leverages Ribo-seq or transcriptomic data integrated with protein-protein interaction networks to accurately predict proteome-wide protein abundances, effectively overcoming the limitations of incomplete proteomics measurements in both bulk and single-cell settings.

Original authors: Jun Ding, Luying Su, Bowen Zhao, Kailu Song, Dingfeng Zou, Jingtao Wang, Ziyang Wang, Wei Liu, Xiuhui Yang, Qinghui Duan, Qi Geng, Rong Huang, Heyun Guo, Yan Lu, Kai Li, David Eidelman, Wei Song

Published 2026-08-06
📖 4 min read☕ Coffee break read

Original authors: Jun Ding, Luying Su, Bowen Zhao, Kailu Song, Dingfeng Zou, Jingtao Wang, Ziyang Wang, Wei Liu, Xiuhui Yang, Qinghui Duan, Qi Geng, Rong Huang, Heyun Guo, Yan Lu, Kai Li, David Eidelman, Wei Song

Original paper licensed under CC BY 4.0 (https://creativecommons.org/licenses/by/4.0/). ⚕️ This is an AI-generated explanation of a preprint that has not been peer-reviewed. It is not medical advice. Do not make health decisions based on this content. Read full disclaimer

Imagine your body is a bustling, high-tech city. In this city, the blueprints for every building are stored in a master library called the DNA. But blueprints alone don't build anything; they need to be copied and sent out to construction sites. These copies are called RNA, and they act like the foremen's orders. Finally, the actual buildings—proteins—are constructed based on those orders. Proteins are the real workers: they build muscles, fight infections, and keep your cells running.

For a long time, scientists had a hard time counting the workers. They could easily read the blueprints (DNA) and the foremen's orders (RNA), but counting the actual buildings (proteins) was like trying to count every single brick in a city while it was being built. The tools were slow, expensive, and often missed a huge chunk of the workforce, especially in tiny, single-cell neighborhoods. This left a big gap: we knew what the instructions said, but we didn't know exactly what was actually getting built. The big question was, could we use the foremen's orders to predict the number of buildings, even if we couldn't see the buildings themselves?

Enter RECIPE, a new digital tool created by a team of scientists that acts like a super-smart detective for this city. The researchers realized that while we can't always see the finished buildings, we can watch the construction crews in action. They used a technique called Ribo-seq, which is like taking a high-speed photo of the construction sites to see exactly where the workers are pausing and how fast they are building. But just looking at the construction speed isn't enough, because sometimes a building gets torn down or modified after it's built.

To solve this, RECIPE uses a "graph neural network," which is a fancy way of saying it maps out how all the proteins know each other. Think of it like a social network for proteins. If you know that Protein A is best friends with Protein B, and you know how many Protein Bs are being made, you can make a very good guess about how many Protein As there should be, even if you can't see Protein A directly. RECIPE combines the construction speed photos (Ribo-seq) with this social network map to predict the total population of proteins in the cell.

The team tested this tool on human and mouse cells, specifically looking at what happens when they disrupted a key part of the construction machinery. They found that RECIPE was incredibly accurate. It didn't just guess the proteins that were easy to see; it successfully predicted the abundance of proteins that were completely invisible to traditional measurement tools. In fact, when they compared RECIPE's predictions to the actual protein counts, the numbers matched up with a correlation of 0.9682 in human cells and 0.9657 in mouse cells. That is a very strong match, suggesting the tool is doing something right.

Perhaps even more impressively, the tool worked in "single-cell" mode. Usually, looking at just one cell is like trying to guess the weather by looking at a single raindrop—it's too small and messy. But RECIPE managed to predict protein levels in individual cells by learning from the "bulk" (large groups) of cells first and then zooming in. It even worked when they only had the blueprint copies (RNA) and no construction photos (Ribo-seq), though it was slightly less accurate in that mode.

The researchers also discovered something fascinating about the proteins that were hardest to predict. These weren't random errors; they were specific types of proteins, often involved in building the cell's skeleton or assembling complex machines. The tool struggled with them because their final numbers are controlled by rules after construction (like how long they last or how they are assembled), not just how fast they are built. By identifying these "hard-to-predict" proteins, RECIPE actually helps scientists find the ones that are being regulated by these hidden, post-construction rules.

In short, RECIPE bridges the gap between the instructions and the final product. It suggests that by combining construction speed data with a map of how proteins interact, we can get a complete picture of the cell's workforce, even for the workers we can't directly see. This doesn't mean we can stop measuring proteins entirely, but it offers a powerful new way to fill in the blanks, especially in situations where direct measurement is impossible or too difficult. The authors show that this approach works well across different species and conditions, providing a more complete view of how our cells function than we had before.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →