Skip to content

Author

Yuhao Zhang

We have 3 of 8 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Review Oct 2026

SkillScriptBench: Benchmarking Self-Evolution of Executable Agent Skill Packages Beyond Markdown

Executable Agent Skills combine natural-language instructions and scripts into reusable packages for LLM agents, and revising them requires fixing errors without breaking correct behavior. Existing benchmarks do not systematically distinguish documentation repair, script repair, and preservation when evaluating skill s...

Yuxuan Liu, Hao-Ran Li, Yu-Hao Zhang et al. · 0 citations

SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision

Evaluated across three main benchmarks, two domain-specific studies, and six LLMs, SkillRevise substantially outperforms one-shot baselines, and the revised skills transfer across both executors and task environments, suggesting that SkillRevise captures reusable procedural knowledge beyond any single executor.

Yuxuan Liu, Zhao-Chen Su, Lin Xie et al. · 17 citations
Preprint Jul 2026

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds

Overall, persistent skill self-evolution is better understood as sparse, validation-filtered search with model- and benchmark-dependent returns, rather than steady improvement from additional rounds.

Yuxuan Liu, Zhaochen Su, Yuhao Zhang et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.