Skip to content
Preprint

Quantifying Risk Under Evolving Uncertainty: Belief-Dependent Robustness for Safe Sequential Decision Making

Aug 2026 · 0 citations · 17 references
Computer Science Engineering

TL;DR

RATTL targets runtime safety for agents, including LLM-based systems, acting under uncertainty, and proves a Safety Sandwich: the RATTL value lies between the uninformed robust value and the full- knowledge optimum, with a gap that vanishes as the posterior concentrates.

Abstract

How cautious should an agent be while it is still learning its environment? We propose RATTL (Risk-Adversarial Total-Reward Learning), which ties caution to epistemic uncertainty: the agent holds a Bayesian posterior over unknown dynamics and plans against a Wasserstein ambiguity set whose radius is a monotone function of that posterior. The radius contracts with evidence, so behaviour interpolates continuously between worst-case robustness and risk-neutral total-reward maximization. The design follows the duality underlying the Entropic Value-at-Risk, which converts the choice of a risk level into the choice of an ambiguity radius. We show the resulting planning problem is well posed under transience and compactness conditions, and prove a Safety Sandwich: the RATTL value lies between the uninformed robust value and the full- knowledge optimum, with a gap that vanishes as the posterior concentrates. In a canonical binary-hazard instance, the induced criterion reduces to Conditional Value-at-Risk at a level set by the posterior entropy. A worked example shows the agent deferring the efficient action until a sharp identification threshold. RATTL targets runtime safety for agents, including LLM-based systems, acting under uncertainty.

View source

Similar papers

#artificial intelligence Preprint Aug 2026

Robust Risk Under Evolving Uncertainty: A Wasserstein Counterpart of the Entropic Value-at-Risk

An agent still learning its environment should be cautious while ignorant and bold once confident. The entropic value-at-risk captures this through a robust-optimization identity---a confidence level fixes the radius of a relative-entropy ball of alternative models---but that ball cannot reach catastrophes the nominal...

D. Ganguly, Jan Křetinský · 0 citations
Sep 2026

Continuous and Monotone Bayesian Nash Equilibrium with Incomplete Information about Player’s Risk Preferences

Abstract. In this paper, we consider a noncollaborative game where each player faces two types of uncertainty: aleatoric uncertainty arising from inherent randomness of underlying data in its own decision-making problem and epistemic uncertainty arising from lack of knowledge and statistical information on the rivals’...

Unknown authors · 4 citations · ⚡1
Open access Aug 2026

A Bayesian composite risk approach for stochastic optimal control and Markov decision processes

The new modeling paradigm subsumes several classical SOC/MDP formulations, including risk-averse and distributionally robust SOC/MDPs as well as partially observed and Bayes-adaptive MDPs, and generates so-called preference robust SOC/MDP models.

Wentao Ma, Zhi-Ping Chen, Huifu Xu · 1 citation
Jul 2026

The Degree of Strategy-Proofness for Risk-Averse Committee Selection

The classic notion of strategyproofness implicitly assumes that a manipulating agent either possesses complete knowledge of what all other agents are going to report, or is willing to take the risk and act as if they know these reports. To capture the profound uncertainty of real-world voters, recent work introduced \e...

Dael Sinay, Rica Gonen · 1 citation · ⚡1
Preprint Aug 2026

Knowing When to Ask for Help: Bayesian Self-Escalation in Hierarchical LLM Agents

The myopic escalation threshold is derived in closed form, characterise the optimal policy via dynamic programming, and it is proved that the optimal policy is a time-varying threshold with no shape assumption on the raw signal.

Nadeem Shaikh · 1 citation

Subjective Risk Decomposition: A New View for Uncertainty Quantification

We present a novel viewpoint for uncertainty quantification. Uncertainty measures are not primitives, in need of axioms and argumentation, but instead consequences, of higher-level modelling decisions. We show how epistemic and aleatoric uncertainty measures can be derived via decomposition of a subjective risk, based...

R. Alamri, Michele Caprio, Gavin Brown · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.