Recent efforts to scale tool-use post-training have largely centered on the synthesis of executable environments, which constitute only one component of a broader agentic interaction system comprising the environment, task, agent harness, and evaluator. Scaling environments in isolation, however, does not guarantee com...
Bo Mao, Hang He, Lin-Ting Wang et al.· 0 citations
Human-written agent skills encode rich workflows for real-world problem solving, but are typically used as external inference-time instructions rather than internalized as reusable model capabilities. We introduce \texttt{SkillGym}, a framework that transforms these skills into executable, verifiable training environme...
Zhi-Long Ge, Yu-Ting Shao, Yu-Tao Yang et al.· 0 citations
This work proposes \textsc{AutoMem}, a text-gradient recursive self-improvement framework for task-adaptive memory architecture search that consistently discovers task-adaptive memory architectures that outperform the strongest human-designed memory baselines.
Lin-Ge Du, Jie Zhou, Yuxuan Cai et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.