Skip to content

Author

Qika Lin

We have 2 of 59 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

R$^2$ Flow: Recursive Self-Improvement via Recursive Skill Evolution

R$^2$ Flow is introduced, a recursive self-improvement framework that alternates policy learning, independent verification, and versioned skill-library updates on a shared-state orchestration graph that improves task accuracy and library-edit precision over heuristic orchestration, reinforcement learning, and skill-evo...

Ming-Da Zhang, Qian-Shuo Huang, Yan-Jin Li et al. · 0 citations
Preprint Aug 2026

Beyond the Stability-Exploration Dilemma: Environmental Regularization for LLM Policy Optimization

Environment-Regularized Policy Optimization (ERPO) replaces the standard Policy-KL regularizer while achieving effective control over query distribution drift, delivering stronger accuracy and substantially more stable behavior under high-temperature decoding and long-horizon training.

Xianlei Zhou, Xiangdi Meng, Yu He et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.