Skip to content

Author

Ying Wen

We have 1 of 9 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

2025

STAR: Efficient Preference-based Reinforcement Learning via Dual Regularization

STAR is proposed, an efficient PbRL method that integrates preference margin regularization and policy regularization that improves feedback efficiency and facilitates more robust reward and value function learning.

Fengshuo Bai, Rui Zhao, Hongming Zhang et al. · 5 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.