Skip to content
Preprint

Tree-of-Experience: Hierarchical Experience Management for Self-Evolving Agents

Aug 2026 · 0 citations · 15 references
Computer Science

TL;DR

Experimental results show that ToE substantially improves both problem-solving performance and efficiency and organizes the experience into a shared tree of analytical perspectives and reasoning paths, whose reliability is calibrated through environmental outcomes to support systematic updating, transfer, and efficient retrieval.

Abstract

Continual self-evolution requires LLM agents to transform environmental interactions into reliable and reusable experience. Existing methods typically refine individual trajectories or abstract shared knowledge from related trajectories, but their experience representations are often disconnected from the underlying reasoning process. This limits feedback attribution, cross-task transfer, and update and retrieval efficiency, particularly in complex reasoning tasks with outcome-level feedback. To overcome this limitation, we propose \textbf{T}ree-\textbf{o}f-\textbf{E}xperience (ToE), a structured experience-management framework that aligns experience organization with the hierarchical reasoning process of LLM agents. Specifically, ToE organizes the experience into a shared tree of analytical perspectives and reasoning paths, whose reliability is calibrated through environmental outcomes to support systematic updating, transfer, and efficient retrieval. The experimental results on \textsc{Game of 24} and \textsc{FinEvolveBench} show that ToE substantially improves both problem-solving performance and efficiency. On \textsc{Game of 24}, ToE achieves a 31.4\% relative improvement in accuracy over the experience-free ToT baseline. On \textsc{FinEvolveBench}, ToE improves tsIC by an average of 41.24\% over the experience-free pipeline across 12 evaluation settings, whereas conventional experience-management methods often underperform experience-free baselines.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

From Interaction Traces to Persistent Skills: Online Evolution for Computer-Use Agents

This work presents an online skill-evolution framework that converts interaction trajectories and evaluator feedback into a persistent, versioned library of reusable procedures and characterize evolving skill libraries as auditable, shared procedural memory that can improve a fixed computer-use stack.

Long-Tao Hu, Xiao Liang, Lin-Chao Zhu · 1 citation

BizSage: A Self-Evolving Multi-Agent Framework for Business Research with Efficient Knowledge Retrieval

BizSage is presented, a multi-agent framework combining corpus-level fine-grained retrieval with quality-driven self-evolution that paves the way for reliable research assistance in economics, business, and the broader social sciences.

Yu-He Wu, Guang-Yu Wang, Jia-Xin Liu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SimSkill: A Self-Evolving LLM Agent for Skill and Knowledge Accumulation in Traffic Simulation

Cumulative culture enables humans to preserve, reuse, and extend knowledge and skills across experiences and generations. Inspired by this principle, we introduce \textit{SimSkill}, a self-evolving agent built around the Simulation of Urban MObility (SUMO) traffic simulator. SimSkill continually identifies capability g...

Qi Liu, Qin-Zheng Wang, Can Li et al. · 0 citations
#human-computer interacti... Preprint Sep 2026

Recursive Organization Improvement: A Modeling Specification for Human--Agent Organizations

Stronger AI agents do not automatically produce better organizations: teams must also learn which work arrangements to retain and when to reconsider them. We propose a modeling specification for recursive organization improvement and evaluate it through an executable checker, a public-record mapping, and controlled sim...

Zi-Long Wang · 0 citations
#artificial intelligence Preprint Sep 2026

MASkills: Continual Skills Optimization for Multi-Agent LLM Systems

MASkills presents a new agent-optimization pipeline that integrates skill-conditioned credit assignment, hierarchical credit aggregation, and momentum-smoothed optimization, enabling agent skill libraries to evolve through refinement, induction, consolidation, and pruning.

Huaiyuan Yao, Xiaoou Liu, Charles Fleming et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.