Skip to content
Preprint

SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure

Aug 2026 · 2 citations · 45 references
Computer Science

TL;DR

SkillZip is presented, an evaluation-free method that compresses a skill by finding its shortest faithful structural explanation, subject to a hard coverage constraint for every extracted trigger, workflow edge, tool requirement, obligation, and output field.

Abstract

Self-evolving agents accumulate reusable skills by appending successful procedures and failure fixes. Over time, the same requirement is often restated in several branches, examples, and warnings, while common action sequences are copied rather than reused. The resulting skill becomes expensive to inject and difficult to maintain. Generic prompt compression is ill-suited to this setting because a skill is not a flat passage: its name and description define when it applies, its workflow controls execution, its tool and output contracts constrain validity, and rare exceptions may remain essential even when no sampled task activates them. Evaluation-guided compression can test these behaviors, but it introduces rollouts, cost, and dependence on the compression-time evaluation set. We present SkillZip, an evaluation-free method that compresses a skill by finding its shortest faithful structural explanation. The intuition is explain once, reference many: state a repeated rule once at the scope where it applies, factor a repeated action sequence into a shared procedure, and keep only the differences as explicit exceptions. We formalize this intuition as a typed minimum description-length objective over a skill contract and a residual, subject to a hard coverage constraint for every extracted trigger, workflow edge, tool requirement, obligation, and output field. The formulation provides simple sharing thresholds, preserves unique rare rules by construction, and supports efficient local updates. SkillZip has a one-shot mode with one structured extraction call and deterministic optimization, and a continual Zip-on-Write mode that integrates each self-evolution patch without replaying tasks or reparsing the full history. Through comprehensive experimental evaluations, we demonstrate the effectiveness and superiority of SkillZip in compression performance, generalizability, and cost overhead.

View source

Similar papers

#artificial intelligence Preprint Aug 2026

SkillZip Pro: Execution-Aware Dynamic Compression of Progressively Loaded Skills for Self-Evolving Agents

This work introduces \method, an evaluation-free compressor for complete, progressively loaded skill bundles, which leaves the agent harness unchanged and emits an ordinary directory and preserves routing, so every required file and directly callable entry remains reachable after rewriting.

Xiaofan Bai, Chao Liu, Hong-Qiang Lin et al. · 0 citations
Preprint Aug 2026

SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries

SkillZip is proposed, an execution-aware procedural abstraction framework that performs contract-preserving compression over section-level graphs that hydrates a compact, dependency-closed context and expands macros only when required.

Xingyu Tan, Xiaoyang Wang, Qing Liu et al. · 2 citations
#artificial intelligence Preprint Sep 2026

SkillGLoW: Procedural-Family Skill Consolidation for Self-Improving Agents on Long-Horizon Task Streams

This work argues the missing unit of reuse is the solving procedure shared by a cluster of related tasks, and builds SkillGLoW (Global-Local Weave) around it: the local skills a task writes from its own execution are aggregated into procedural families and compressed into de-instantiated global priors.

Ao Yan, Xin Zhang, Jiawei Du et al. · 0 citations
Preprint Aug 2026

Practice Makes Unsafe: Skill Misevolution in Self-Improving LLM Agents

Self-improving LLM agents convert successful trajectories into persistent cross-task state. An unsafe success can thereby become reusable policy after its triggering input disappears. Skill evolution makes this failure measurable by distilling operational trajectories into executable, transferable, and inspectable proc...

Xutao Mao, Liang Zhao, Xiang Zheng et al. · 0 citations
Preprint Aug 2026

Learning Globally Reusable Skills for Coding Agents

This work proposes GSE, a globalized skill evolution framework that jointly optimizes skill compatibility and skill generalization, and maintains a Skill Relation Graph (SRG) that explicitly models and co-evolves inter-skill relationships.

Chen Yang, Jiashuo Tian, Zi-Qi Wang et al. · 2 citations · ⚡1
Preprint Aug 2026

JailbreakSkill: Scaling Automated Red-Teaming with Reusable and Ever-Evolving Skills

This work introduces \textsc{JailbreakSkill}, a skill-centric framework for scaling automated red-teaming through reusable and continuously evolving attack capabilities, which packages existing attack strategies into modular, agent-ready skills that can be directly reused and adaptively selected across tasks and target...

Xiaoyu Wen, Jiajia Li, Zhida He et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.