Skip to content

Author

Bingxuan Li

We have 6 of 22 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

WEFT: Scaling Tool-Use Post-Training for General-Purpose Agents

Recent efforts to scale tool-use post-training have largely centered on the synthesis of executable environments, which constitute only one component of a broader agentic interaction system comprising the environment, task, agent harness, and evaluator. Scaling environments in isolation, however, does not guarantee com...

Bo Mao, Hang He, Lin-Ting Wang et al. · 0 citations

Useful Memories Become Faulty When Continuously Updated by LLMs

This work traces the regression to the consolidation step rather than the underlying experience: the same trajectories yield qualitatively different memories under different update schedules, and an episodic-only control that simply retains those trajectories remains competitive with the consolidators the authors test.

Dylan Zhang, Yan-Shan Lin, Zheng Wu et al. · 10 citations
Preprint Aug 2026

ChronoVision: Temporal Reasoning via Latent State Reconstruction

This work proposes ChronoVision, a multimodal framework designed to align visual logic with latent imagery, and introduces Vbvr-VQA, a novel dataset that evaluates temporal tracking by reformulating video reasoning into a strict image-ordering task.

Yi-Fan Shen, Jian Xu, Boyi Li et al. · 1 citation
Preprint Jul 2026

CUADebug: Diagnosing and Repairing Computer-Use Agent Failures

Results show that CUA root-cause diagnosis can provide actionable repair signals rather than merely post-hoc explanations and show that CUA root-cause diagnosis can provide actionable repair signals.

Wei-Jia Zhang, Kunlun Zhu, Ze-Yi Liu et al. · 1 citation
Preprint Aug 2026

EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents

EvoHarness-RL is introduced, which exposes Belief, Progress, and Experience (BPE) as policy-facing harness state and reveals two key dynamics: harness annealing, where training internalizes recurring harness-use patterns into the model policy and shifts the agent from frequent harness calls toward selective external-st...

Xuying Ning, Dongqi Fu, Tianxin Wei et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.