Skip to content

Author

Yong-Jie Guan

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Bandits with Probing: Optimal Regret and the Limits of Winner Feedback

A learner probes at most $k$ of $n$ arms each round, receives the maximum of their rewards in $[0,1]$, and competes with the best fixed arm. When does the probing advantage pay for learning? We determine two minimax laws. Under independent stochastic rewards with winner feedback (the maximum and a winning label), or on...

Yong-Jie Guan · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.