Skip to content

Author

Yanxiao Zhao

University of Chinese Academy of Sciences

We have 2 of 8 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Jul 2026

RLLMNav: boosting object goal navigation through seamless integration of historical experience with large language model

This work proposes RLLMNav, a method that integrates historical experience of RL with commonsense reasoning of LLM through a confidence-gated routing mechanism, and utilizes commonsense knowledge extracted from an LLM to suggest frontiers.

Kexun Chen, Qianlei Wang, Zhixuan Shen et al. · 0 citations
Jul 2026

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL

This work introduces ScaleCUA, a unified framework that scales online RL for CUAs via verifiable task synthesis and efficient training, and designs VeriGen, an end-to-end framework for generating verifiable RL tasks through iterative docker interactions and a multi-agent feedback loop.

Bowen Lv, Xiao Liu, Yanyu Ren et al. · 4 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.