Skip to content

Author

Zhiyuan Liu

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Conference Feb 2025

AgentRM: Enhancing Agent Generalization with Reward Modeling

This work finds that finetuning a reward model to guide the policy model is more robust than directly finetuning the policy model, and proposes AgentRM, a generalizable reward model, to guide the policy model for effective test-time search.

Yu Xia, Jing-Ru Fan, Weize Chen et al. · 26 citations · ⚡3

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.