Skip to content

Author

Shuaiqiang Wang

We have 6 of 124 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Navigating Sparse Evidence: Agentic Visual RAG via Explicit Context Selection and Consolidation

This work proposes SCoRE (Selection and Consolidation for Robust Evidence), a unified agent loop for explicit evidence selection and consolidation, which decouples final reasoning from exploratory trial-and-error while ensuring strict visual grounding via indexed claim-to-image linkages.

Yucheng Shen, Lingyong Yan, Jiulong Wu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SWRouter: Similarity-Contractive Window Routing for Multi-Turn Large Language Model Conversations

This paper proposes SWRouter, a Similarity-Contractive Window Router for multi-turn large language model routing that combines a similarity-based context segmentation mechanism for prompt construction with a dual-metric evaluation framework that decouples construction accuracy from router performance.

Yu Wang, Yuchen Li, Rui Kong et al. · 0 citations
#artificial intelligence Preprint Feb 2026

Not All Preferences Deserve Gradients: Understanding Gradient Utility in Offline Reasoning Alignment

SAGE (Stability-Aware Gradient Efficiency), which maintains difficulty-stratified candidate pools refreshed during training and selects pairs within each pool by a forward-pass signal-to-curvature score, outperforms full-data and size-matched baselines while producing substantially smoother optimization trajectories.

Hui Wu, Hengyi Cai, Jin-Man Zhao et al. · 0 citations
Jul 2026

DocOps: A Verifiable Benchmark for Autonomous Agents in Complex Document Operations

DocOps is introduced, a deterministically verifiable evaluation framework underpinned by a hierarchical taxonomy that deconstructs document operations inspired by real-world practices into atomic dimensions and escalating workflow complexities that exposes the capability boundaries of agents in maintaining global docum...

Jiazhen Jiang, Boxi Cao, Lingyong Yan et al. · 1 citation
Preprint Aug 2026

DuMateBench: Evaluating Autonomous Agents in Complex Real-World Workflows

DuMateBench, a real-session benchmark reconstructed from anonymized and privacy-screened user sessions collected from a large-scale production agent platform, is introduced, showing that performance under environmental perturbations is jointly shaped by the capabilities of the LLM and the surrounding agent framework.

Zechun Niu, Yu-Kun Zhao, Jia-Xin Zhang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.