Skip to content

Author

N. Kuang

We have 4 of 16 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Review Sep 2026

ArcticSwarm: Deferring Early Consensus in Long-Horizon Multi-Agent Research

The results show that restricting peer reads during evidence gathering and strengthening commitment boundaries before a hypothesis is shared can broaden search and improve long-horizon multi-agent deep research.

Soyoung Yoon, Bo-Yi Liu, Yi-Te Wang et al. · 0 citations
Jul 2026

RRPO: Reference-Relative Policy Optimization with Stratified Conditional Rollouts

Group Relative Policy Optimization (GRPO) has shown strong effectiveness in reinforcement learning from verifiable feedback, where sampled rollouts can be compared within a group using task-provided correctness signals. However, extending group-relative optimization beyond verifiable settings is challenging because suc...

Yuxin Xiong, Xun-Yi Jiang, Rohan Surana et al. · 0 citations
Preprint Aug 2026

Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning

Recoverability-Aware Intervention Learning (RAIL), a training-time framework that learns how to generate rollouts based on the improvement produced by each intervention, consistently improves performance under limited rollout budgets.

Zheyuan Zhang, Man-Qing Mao, Hong Wang et al. · 0 citations
Jul 2026

Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows

This work introduces Spider 2.0-AIFunc, a benchmark of 465 verified instances across 125 real-world databases covering six types of AI functions on the Snowflake platform, and finds that the strongest proprietary models reach 67-70% execution accuracy while the best open-source model achieves 58.1%, a gap driven primar...

Tianyang Liu, Canwen Xu, Fangyu Lei et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.