Skip to content

Author

Guoxi Zhang

5 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Beyond Forgetting: Diagnosing and Harnessing Shared Reasoning in Continual RLVR

Reinforcement learning with verifiable rewards (RLVR) commonly post-trains reasoning models on multiple tasks, while rerunning multitask RLVR (MTRL) as new tasks are added makes capability expansion costly. We therefore study continual RLVR, which updates the existing model as each task arrives. The central question is...

Li-Rui Luo, Guo-Xi Zhang, Hong-Ming Xu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SkillRubric: Co-Evolving Actor Guidance and Evaluator Rubrics for Multimodal Agents

Recent work incorporates reusable skills distilled from past interactions into multimodal agent training, providing procedural guidance for long-horizon planning and tool use. However, policy optimization in these methods remains driven primarily by sparse outcome rewards, providing little supervision for intermediate...

Bing Jiang, Guo-Xi Zhang, Jasper Wang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

AgentBoundary: Counterfactual Evaluation of Safety in Tool-Using LLM Agents

Safety alignment for large language models (LLMs) in conversational settings is largely framed around whether to answer or refuse a request. In agentic settings, however, the same models must decide whether to act as permission-critical evidence emerges during execution. This creates a distinct challenge: apparent risk...

Tian-Zhuo Yang, Zi-Rui Mi, Yan-Tao Huang et al. · 0 citations
#machine learning Preprint Aug 2026

Continual Reasoning Gym: Diagnosing and Harnessing Shared Reasoning in Continual RLVR

Reinforcement learning with verifiable rewards (RLVR) commonly post-trains reasoning models on multiple tasks, while rerunning multitask RLVR (MTRL) as new tasks are added makes capability expansion costly. We therefore study continual RLVR, which updates the existing model as each task arrives. The central question is...

Lirui Luo, Guo-Xi Zhang, Hong-Ming Xu et al. · 0 citations
Preprint Jul 2026

UniLM-Nav: A Unified Framework for Zero-Shot Last-Mile Navigation

Mobile manipulation requires a robot to navigate to a target object or receptacle and then perform intended manipulation. However, reaching the vicinity of the target does not guarantee a manipulation-ready base pose, a problem known as last-mile navigation. Prior methods for last-mile navigation either rely on manual...

Zhuofan Zhang, Tianxu Wang, Guoxi Zhang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.