Beyond Optimal Rates in Stochastic Optimization: Trajectory-Adaptive Stopping Rules
This work treats the evolving SGD trajectory as a sequential experiment whose observations provide evidence about the unknown optimization error, and develops new recursive confidence-sequence techniques and a general time-uniform empirical Bernstein inequality for adapted processes with time-varying conditional means...