Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Sep 2026

Learning from Teacher Continuations at Student States

We present OLIVE (OnLine InterVEntion). At each iteration, the evolving student policy generates a new prefix, the teacher continues it autoregressively, and the student is updated using cross-entropy computed on the teacher-generated tokens. Each design choice targets a corresponding limitation of existing distillatio...

Hao-Jin Wang, Dylan Zhang, Huai-Bo Chen et al. · 0 citations

Useful Memories Become Faulty When Continuously Updated by LLMs

This work traces the regression to the consolidation step rather than the underlying experience: the same trajectories yield qualitatively different memories under different update schedules, and an episodic-only control that simply retains those trajectories remains competitive with the consolidators the authors test.

Dylan Zhang, Yan-Shan Lin, Zheng Wu et al. · 10 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.