On-policy self-distillation uses a model as its own teacher to provide dense supervision for reasoning, often through reference-solution conditioning. Providing privileged information does not by itself ensure effective token-level supervision throughout long responses. We introduce Activation-Conditioned Self-Distilla...
Zhe-Xi Lu, Subhajit Chaudhury, Tejaswini Pedapati et al.· 0 citations
Large language model (LLM) agents increasingly operate over long-horizon interactions involving tool use, persistent state, evolving authorization, and external environment feedback. In such settings, safety failures may emerge only after multiple turns, yet existing evaluations often reduce agent behavior to task or a...
Sadia Asif, Mohammad Mohammadi Amiri, Momin Abbas et al.· 0 citations
The architecture family implementing this function class is named CoFrGeNets - Continued Fraction Generative Networks, and novel architectural components based on this function class that can replace Multi-head Attention and Feed-Forward Networks in Transformer blocks while requiring much fewer parameters are designed.
Amit Dhurandhar, Vijil Chenthamarakshan, Dennis Wei et al.· arXiv.org· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.