Skip to content
Preprint

Statistical Properties of Robust Learning under Distributional Shifts

Aug 2026 · 0 citations
Mathematics Computer Science

TL;DR

Finite-sample generalization error bounds in the shifted target environment for both DRO and RS are derived, which fill a gap in understanding the statistical properties of robust learning methods under distributional shifts and provide a principled basis for comparing DRO and RS.

Abstract

Distributional shifts arise when the target deployment environment differs from the source environment that generated the training data. Robust learning frameworks such as Distributionally Robust Optimization (DRO) and Robust Satisficing (RS) aim to address this challenge, yet their finite-sample guarantees under such shifts, and their systematic comparison, remain underexplored: existing analyses typically establish guarantees either in the source environment or for adversarial worst-case performance over an ambiguity set. This paper instead studies generalization error in the target environment---the excess loss under the shifted target distribution. Our contributions are threefold. First, we derive finite-sample generalization error bounds in the shifted target environment for both DRO and RS. These bounds explicitly characterize the trade-off between reduced sensitivity to shift and the regularization penalty induced by each method's robustness hyperparameter, and they avoid the curse of dimensionality associated with Wasserstein empirical concentration. Second, when partial shift information such as shift magnitude or direction is available, we propose information-directed hyperparameter calibrations and compare the two methods given the same information. Under these calibrations, and in the partial-information regimes we study, DRO and RS exhibit complementary theoretical and empirical behavior. Finally, we apply the framework to a network lot-sizing problem, using it to interpret how robust policies respond to positive shifts in the demand distribution. Together, these results fill a gap in understanding the statistical properties of robust learning methods under distributional shifts and provide a principled basis for comparing DRO and RS.

View source

Similar papers

#machine learning Preprint Sep 2026

Distributionally robust linear regression through the lens of adversarial training

Distributionally robust optimization (DRO) studies parameter estimation under uncertainty in the underlying probability distribution and has emerged as a principled framework for analyzing robustness and generalization. In particular, Wasserstein DRO, with distributional uncertainty induced by the Wasserstein distance,...

Elis Stefansson, David Vävinggren, Antônio H. Ribeiro · 0 citations
#machine learning Preprint Sep 2026

Domain Adaptation with Target Information via Doubly-Anchored Distributionally Robust Optimization

A doubly-anchored DRO framework whose ambiguity set is the intersection of $\phi$-divergence balls centered at the source law and a source-completed target reference law, the latter pairing the target covariate law with the source conditional law is introduced.

D. Kepplinger, Anand N. Vidyashankar · 0 citations
#machine learning Preprint Sep 2026

Brenier Meets Adversarial Training: Optimal Transport Geometry for Robust Learning

A penalized DRO formulation in which the adversary may choose any distribution but incurs a Wasserstein penalty for deviating from the empirical distribution is studied, showing that the adversary's problem can be reformulated as an optimization problem over transport maps that push empirical samples to adversarial one...

Alireza Abdollahpoorrostam, Ehsan Sharifian, Buse Sen et al. · 0 citations
#artificial intelligence Preprint Sep 2026

General Quantification of Covariate and Concept Shifts

This paper proposes a key notion: $\gamma^{*}\!$-concept shifts, and derive a general error bound unifying covariate and $\gamma^{*}\!$-concept shifts, which applies to broad loss functions, label spaces, and stochastic labeling and develops estimators for these shifts with concentration guarantees.

Hong-Bo Chen, L. Xia · 0 citations
Preprint Sep 2026

Conformal-DRO: Distributionally Robust Optimization with Conformalized Ambiguity Set

Conformal-DRO is proposed, which uses nested conformal regions to construct an ambiguity set for the future latent law, which covers this law with probability at least $1-\alpha$ in finite samples, without estimating underlying latent laws or their mixing mechanism.

Lu-Hao Zhang, Shi-Xiang Zhu · 0 citations
Preprint Aug 2026

Wasserstein Filtering: A Sample Selection Method for Robust Distribution Learning

This work proposes Wasserstein Filtering (WF), a novel sample selection framework that discards a fraction of suspicious samples and estimates the target distribution using the empirical measure of the remaining data, and proves that the WF estimator achieves minimax optimality over distribution families with bounded c...

Yi-Kai Xu, Zhao Chen, Jian Huang · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.