Scaling model-generated data is usually viewed as improving distillation: more examples should increase coverage, reduce noise, and produce stronger students. We show a second effect: larger datasets can make subtle teacher-specific signals easier to detect in the trained student, even when examples are off-task and ne...
Zhichen Dong, Zhi-Xuan Liu, Yuanjiu Fan, et al.· 1 citation
The original robust value frontier, support-wise linear-programming algorithm, and binary-action fractional-knapsack specialization are embedded into this implementation framework and embeds the original robust value frontier, support-wise linear-programming algorithm, and binary-action fractional-knapsack specializati...
Trait-direction drift is proposed and validated as a mechanism for subliminal learning: biased generation creates measurable preference gaps in teacher data, and student-recognizable gaps induce trait-aligned updates during supervised fine-tuning that accumulate into behavioral transfer.
Zhi-Xuan Liu, Zhichen Dong, Yuanjiu Fan, et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.