Skip to content

Author

Jingqi Ye

We have 2 of 7 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Sep 2026

Rethinking Token Reweighting for SFT: Suppress, Reverse, and Extrapolate Learned Features

SCALE (Selective Control of Adaptation via Local Entropy), an entropy-guided adaptation-strength-control method that freezes the pretrained model and the SFT delta and learns bounded token- and module-specific gates by minimizing predictive entropy alone is proposed.

Cun-Chun Li, Hao-Nan He, Yi-Fan Gao et al. · 0 citations
Jul 2026

MoE2-LoRA: When MoE Models Meet MoE-style Low-Rank Adaptation

This paper makes the first attempt to fine-tune MoE models with MoE-style low-rank adaptation via a dual-channel Routing-Conditioned Projection module, which reuses base router activations to inform LoRA routing and introduces a single global LoRA expert pool shared across all layers.

Qingyu Yang, Haonan He, Minglei Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.