Skip to content

Author

Haonan Lu

We have 3 of 34 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Learn from the Gap: Differential-Aware Advantage Pruning with Adaptive Rollout Sampling for GRPO

Recently, Group Relative Policy Optimization (GRPO) and its variants have been developed for policy optimization and demonstrated notable performance gains. However, these methods usually incur substantial computational overhead due to per-question multi-rollout sampling and repeated per-token probability evaluation ac...

Jia-Hua Yang, Zhiwei Yang, Xian-Peng Zhang et al. · 0 citations
Jul 2026

TimeThink: Reasoning with Time for Video LLMs

TimeThink is proposed, a reinforcement learning framework that explicitly guides temporal evidence discovery in Video-LLMs and introduces a step-wise temporal process reward that provides localized credit assignment for these clues and a joint process--outcome optimization objective that balances reasoning fidelity wit...

Handong Li, Longteng Guo, Zikang Liu et al. · 0 citations
Jun 2026

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation

This work starts from an empirical observation: when query-relevant visual evidence is explicitly strengthened using the model's own attention, generation becomes more accurate, suggesting that many failures do not arise solely from missing perception, but from an insufficient tendency to trust the evidence the model h...

Xin Zou, Hao Deng, Yibo Yan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.