Skip to content

Author

Sijia Cui

We have 2 of 9 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

2025

STAR: Efficient Preference-based Reinforcement Learning via Dual Regularization

STAR is proposed, an efficient PbRL method that integrates preference margin regularization and policy regularization that improves feedback efficiency and facilitates more robust reward and value function learning.

Fengshuo Bai, Rui Zhao, Hongming Zhang et al. · 5 citations
Preprint Aug 2026

HyMem: Hierarchical Context Management for Long-Horizon Agents via Information Isolation

HyMem is a hierarchical framework that explicitly separates the agent's context into distinct functional layers to separate high-level planning from execution and complex analysis, allowing the model to maintain focus and accuracy across complex, long-horizon tasks.

Xinqi Wang, Jinwei Xiao, Sijia Cui et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.