← Latest papers
🔬 materials science

Truncated automatic sparse differentiation for machine learning interatomic potentials

This paper proposes truncated automatic sparse differentiation to efficiently compute higher-order derivatives for machine learning interatomic potentials by exploiting physical locality to discard negligible long-range interactions, achieving order-of-magnitude speedups with minimal impact on predicted observables.

Original authors: Marcel F. Langer, Adrian Hill, Michele Ceriotti

Published 2026-09-18
📖 5 min read🧠 Deep dive

Original authors: Marcel F. Langer, Adrian Hill, Michele Ceriotti

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer

Atoms are the fundamental building blocks of matter, but they are rarely still. In a solid piece of metal or a porous crystal, these atoms constantly vibrate, jiggling against one another in a complex, collective rhythm. To understand how a material behaves—how it conducts heat, how it vibrates, or how it might react to stress—scientists must map the invisible forces that bind these atoms together. This map is known as the potential energy surface. For decades, the most accurate way to draw this map has been through quantum mechanics, a set of rules that describe the behavior of electrons and nuclei with extreme precision. However, these calculations are so computationally heavy that they become impossible for large systems, such as the giant, sponge-like crystals used to store gas or purify water.

To overcome this barrier, researchers have turned to machine learning. These artificial intelligence models are trained on the results of those expensive quantum calculations, learning to predict the energy of a system based on the positions of its atoms. Once trained, these models can predict energies and the forces pushing on atoms almost instantly, allowing scientists to simulate the motion of millions of atoms over time. But there is a catch. While these models can easily predict how atoms move, they struggle to predict the higher-order details of their vibrations. Specifically, calculating the "curvature" of the energy landscape—the mathematical description of how the forces change as atoms move—has been considered too slow for anything but the smallest systems. Without this information, scientists cannot accurately predict experimental properties like heat capacity or the specific frequencies at which a material vibrates.

A team of researchers has now found a way to bypass this bottleneck, allowing them to calculate these complex vibrational properties for massive materials that were previously out of reach. Their work focuses on a specific type of machine learning model that mimics how atoms communicate with their neighbors. In these models, information travels from one atom to the next in steps, much like a rumor passing through a crowd. The researchers realized that because physical forces weaken rapidly with distance, the mathematical description of these vibrations is not a solid block of numbers, but a sparse one, filled mostly with zeros. By recognizing this hidden structure, they developed a method to compute the full picture of atomic vibrations by ignoring the zeros and focusing only on the meaningful connections.

The researchers tested this approach on a variety of large, porous materials, including metal-organic frameworks and zeolites, which are used in industrial applications for their ability to trap molecules. They applied their method to several state-of-the-art machine learning models, including ones designed to work across the entire periodic table. The results showed that while the method could compute the exact vibrational properties of these giant systems, the speed improvement over traditional methods was modest. This was because the models they used had been designed to look at a very wide range of neighbors, meaning the "zeros" in the mathematical description were not as numerous as hoped. The interaction range of these models was simply too broad for the sparsity trick to yield massive gains on its own.

However, the team discovered a powerful refinement. They realized that they could intentionally ignore the very weak connections between atoms that are far apart, a process they call truncation. By discarding these tiny, distant interactions, they could reduce the computational cost dramatically. In many cases, this approach made the calculations ten to twenty times faster. Crucially, this speed came with almost no loss in accuracy. The predicted heat capacities and vibrational frequencies remained nearly identical to the exact results, even though the calculation had skipped a significant portion of the data. The errors introduced by ignoring these distant forces were so small that they fell well below the threshold of what matters for predicting real-world experimental outcomes.

This finding suggests that for many practical applications, scientists do not need to calculate every single interaction in a massive system. They only need to account for the strong, local connections that dominate the physics. The researchers demonstrated that by using this truncated approach, they could compute the vibrational properties of a material containing over six thousand atoms in a matter of minutes on a single graphics card. Previously, such a calculation would have been prohibitively slow or impossible. This opens the door to systematically testing machine learning models against real experimental data, rather than just relying on simulations. It allows researchers to fine-tune these models using actual measurements from the lab, potentially leading to more accurate predictions for new materials.

The work also highlights a new design principle for future machine learning models. The researchers found that the depth of the model—how many steps of neighbor-to-neighbor communication it performs—directly controls how far the forces extend. Models that look too far ahead create dense, complex calculations that are hard to speed up. By keeping the receptive field of these models more compact, scientists can ensure that the resulting calculations remain sparse and fast, without sacrificing accuracy. This insight bridges the gap between the theoretical power of machine learning and the practical needs of materials science, turning a theoretical possibility into a routine tool for exploring the physical world.

Ultimately, this research transforms how we can study the vibrations of matter. It moves the ability to predict complex thermal and vibrational properties from the realm of small, simple molecules to the scale of giant, industrial crystals. By leveraging the natural decay of forces in the physical world, the researchers have shown that we can compute the behavior of large systems with high precision and low cost. This capability is essential for the next generation of materials discovery, where understanding how a material handles heat and vibration is just as important as knowing its chemical composition. The path forward is now clear: use these efficient methods to compare machine learning predictions directly with experimental results, refining our understanding of matter one calculation at a time.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →