Skip to content

Author

Jiangnan Yang

We have 7 of 10 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

Pythia: Toward Foundation World Models for Multimodal Time Series

Time-series foundation models offer a unified approach to forecasting across heterogeneous domains. Textual context and auxiliary observations provide complementary information about temporal dynamics, yet reusable multimodal predictive representations remain underexplored. We introduce Pythia, a foundation world model...

Xilin Dai, Hong-Zhou Chen, Yi-Fan Hu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

When Tomorrow Becomes Today: Self-Evolving Policies for Agentic Time-Series Forecasting

Agentic time series forecasting concerns systems whose underlying mechanisms evolve, making the relative effectiveness of numerical models, reasoning strategies, and intervention rules inherently time-varying. Consequently, a time series agent must adapt the forecasts it produces and the orchestration policy that deter...

Yi-Fan Hu, Xilin Dai, Zhi-Yuan Qu et al. · 1 citation
Preprint Aug 2026

ARC: Fair Relative Advantage Comparison in Open-Ended Real-World Interaction

The proposed ARC (Advantage Regularization via Conditioning), a training recipe that restores fairer relative comparison through strategy-conditioned rollout grouping, together with hybrid rewards and entropy regularization, is proposed.

Yong-Qi Tong, Tan Li Hui Faith, C. Marcus et al. · 0 citations
Preprint Aug 2026

Ask, Condition or Abstain: Reinforcement Learning for Missing-Premise Reasoning

ACA-RL supports a new mission for NLP evaluation: measuring whether models can recognize when a task is underdetermined and handle uncertainty, not only whether they can answer fully specified questions.

Yong-Qi Tong, Zhenyu Zhang, Zimou Liu et al. · 1 citation
Preprint Aug 2026

STAGE: Controlled Objective Admission for Multi-Preference LLM Alignment

This work proposes \methodname, a stability-guided active-set controller for controlled objective admission, a stability-guided active-set controller for controlled objective admission in reward-vector RLHF, which positions objective-entry timing as a concrete control variable in reward-vector RLHF.

Yong-Qi Tong, Z. Zhang, Ruirui Wang et al. · 1 citation
Preprint Aug 2026

Diagnosis Before Recovery: Turning Agent Failures into Selective Self-Correction

DARC is proposed, a diagnosis-guided recovery harness that profiles task-family failure modes, prunes mismatched interventions from a shared recovery library, and freezes a verifier-selected success-cost policy for deployment, providing a practical route toward more reliable agents in domains where compiler-like feedba...

Pan Wang, Yihao Hu, Hang Wang et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.