← Latest papers
📊 statistics

Robust Risk Under Evolving Uncertainty: A Wasserstein Counterpart of the Entropic Value-at-Risk

This paper introduces the Wasserstein entropic value-at-risk, a robust risk measure that overcomes the limitations of traditional entropic value-at-risk by using optimal transport balls to account for catastrophic scenarios deemed impossible by nominal models, thereby enabling a dynamic decision-making framework that transitions from caution to confidence as environmental uncertainty decreases.

Original authors: Deep Kumar Ganguly, Jan Křetínský

Published 2026-08-20
📖 5 min read🧠 Deep dive

Original authors: Deep Kumar Ganguly, Jan Křetínský

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer

In the world of autonomous machines, from delivery drones to self-driving cars, there is a fundamental tension between speed and safety. An agent, or machine, must make decisions while it is still learning about its surroundings. When it knows very little, it should be extremely cautious, avoiding any path that might lead to disaster. But once it has gathered enough evidence to understand the environment, it should become bold enough to act efficiently. The challenge lies in creating a system that can smoothly transition from extreme caution to confident action as its knowledge grows, without ever making a fatal mistake during the learning process. This requires a way to measure risk that changes dynamically: the machine's own uncertainty about the world should dictate how much it fears the unknown.

Researchers at the Technical University of Munich and Masaryk University have developed a new method to solve this problem, offering a safer way for machines to navigate uncertain environments. They focused on a specific type of risk calculation used to protect against worst-case scenarios. Previous methods relied on a mathematical concept that, while useful, had a critical flaw: if a machine's current model of the world said a disaster was impossible, the safety system would completely ignore that disaster, even if it was physically possible. The new approach replaces this old method with a different mathematical tool that allows the machine to consider catastrophic outcomes that its current model deems unlikely, but not impossible. This ensures that the machine remains vigilant against rare but devastating events until it is absolutely certain they cannot happen.

The researchers built their solution around a concept called the "entropic value-at-risk," which is a standard way to calculate the worst likely loss a system might face. This standard method works by imagining a circle of possible alternative realities around the machine's current best guess. If the machine is unsure, the circle is large, and it considers many different possibilities. As the machine learns, the circle shrinks. However, the shape of this circle in the old method was rigid. It was constructed using a rule that meant if the machine's best guess assigned zero probability to a specific disaster, no amount of uncertainty could ever bring that disaster back into consideration. The disaster was effectively invisible to the safety system, creating a blind spot that could lead to fatal errors when the machine was overconfident.

To fix this, the team introduced a new measure they call the "Wasserstein entropic value-at-risk." Instead of using the rigid circle of the old method, they used a different geometric shape based on the physical distance between outcomes. Imagine a machine trying to cross a canyon. If the wind is calm, the machine might guess the air is still. If a sudden gust hits, the machine needs to know that a storm is possible, even if its current data says the air is calm. The old method would say, "Since my data says calm, a storm is impossible, so I will ignore it." The new method says, "A storm is far away from my current guess, but it is reachable. I will pay a price to consider it, proportional to how far away it is." This allows the safety system to account for disasters that are physically possible but currently seem unlikely, ensuring the machine hedges against them until it is sure they are not happening.

The researchers proved that this new method is mathematically sound and fits perfectly into the existing hierarchy of risk measures. They showed that it behaves exactly as a safe system should: it is conservative when the machine is ignorant and becomes less conservative as the machine learns. Crucially, they demonstrated that this new measure strictly accounts for those "zero-probability" catastrophes that the old method ignored. In their tests, when they introduced a disaster scenario that the machine's model said had zero chance of occurring, the old method's risk calculation did not change at all, no matter how dangerous the scenario became. The new method, however, immediately increased its caution level, reflecting the true danger of the situation.

To make this practical for real-time decision-making, the team linked the size of the safety margin directly to the machine's own "belief entropy," a measure of how confused or uncertain the machine is. When the machine is confused, the safety margin is wide, forcing it to act slowly and carefully. As the machine gathers more data and its belief sharpens, the margin automatically shrinks, allowing it to move faster. This creates a closed-loop system where the machine's caution evolves naturally with its knowledge. They tested this on a simulated drone crossing a canyon. The drone started by hovering and creeping forward, the safest but slowest option. As it gathered evidence that the wind was calm, its internal uncertainty dropped, the safety margin tightened, and it switched to a fast cruising policy. The transition happened automatically, without a hard-coded switch, and the system maintained a guaranteed safety bound throughout the entire process.

The results of their simulations confirmed that the new method works exactly as the theory predicted. The machine successfully avoided catastrophic failures that the old method would have missed, while still being able to move efficiently once it was confident. The researchers verified that their mathematical formulas matched the results of complex computer simulations to a very high degree of precision. They also showed that the system converges quickly to a stable solution, meaning it can be used in real-time applications. By replacing a flawed geometric assumption with a more robust one, the team has provided a way for autonomous agents to be both safe and efficient, ensuring they remain cautious when they should be and bold when it is safe to be so. This work offers a concrete path forward for building machines that can learn and adapt in the real world without sacrificing safety for speed.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →