Skip to content

Author

Zu-Jie Wen

We have 2 of 16 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Robust Code RL via Faulty-Code-Driven Test case Synthesis and Dense Reward Shaping

Experimental results show that RL fine-tuning of Qwen3-32B via RobustTests achieves a 3% absolute gain on LiveCodeBench, demonstrating its effectiveness in advancing LLM code generation proficiency.

Yiwen Zhang, Xiaodong Yan, Zhenyu Huang et al. · 0 citations
Jul 2026

Ring-Zero: Scaling Zero RL to a Trillion Parameters for Emergent Reasoning

This work presents a stable and efficient training pipeline, incorporating algorithmic and system optimizations such as clipped importance sampling, training-inference ratio correction, and mixed-precision control, and proposes a structured evaluation framework across three dimensions: comprehensibility, reproducibility, and efficiency.

Xinyu Tang, Gangqiang Cao, Yurou Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.