The Qinghai-Tibet Plateau, a climate-vulnerable source of Asia’s major rivers, harbors underexplored viral communities critical to ecosystem functions. By integrating 597 metagenomes from the Yangtze, Yellow, Lancang, and Yarlung Tsangpo rivers with 85 public available glacial metagenomes (Tibetan Glacier Genome and Ge...
Ying-Hao Li, Tian-Yi Chen, Pengwei Li et al.· Nature Communications· 0 citations
Diagnostics reveal that RL on PLMs is governed by two reward properties: verifiability, whether the reward is a fixed environment or a learned surrogate vulnerable to distribution shift, and coverage, the fraction of sequence space giving an informative gradient.
Hanqun Cao, Hongrui Zhang, Junde Xu et al.· Proceedings of the 32nd ACM...· 0 citations
Reinforcement learning (RL) is increasingly applied to Protein Language Models (PLMs), yet its effectiveness varies across tasks, and standard metrics such as pass@k can rise even when the model's solvable problem set is shrinking. We introduce two capability-level diagnostics. The Expansion-Shrinkage Ratio (ESR) measu...
Hanqun Cao, Hongrui Zhang, Junde Xu et al.· Proceedings of the 32nd ACM...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.