Skip to content

Author

Zengrui Jin

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

FSA-GRPO: Teaching Auditory LLMs to Use Few-shot Demonstrations

This work introduces Few-Shot Aware GRPO (FSA-GRPO), an RL-based post-training recipe that uses a specially designed reward to encourage the model to leverage few-shot demonstrations, thereby strengthening its few-shot adaptation ability.

Haolong Zheng, Siyin Wang, Xulin Fan et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.