Preprint
Aug 2026
RecurSE: Bounded Recursive Self-Evaluation for LLM Rubric Judges
This work eliminates external gold supervision from the RL training reward: the model's own evaluative capability generates learning signals for its optimization -- a closed-loop setting of bounded recursive self-improvement (RSI) termed Recursive Self-Evaluation (RecurSE).
Kai-Yuan Liu, Ziyuan Zhuang, Rongxiang Weng et al.
· 0 citations