PaperScope
LIVE · 2026-09-03 05:40 UTC

MASkills: Continual Skills Optimization for Multi-Agent LLM Systems

Huaiyuan Yao, Xiaoou Liu, Charles Fleming, Tianlong Chen, Hua Wei

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.02094 v1
Category
Submitted
2026-09-02

Abstract

LLM-based multi-agent systems have shown strong performance on complex tasks, yet continual improvement from interaction experience remains challenging. Existing self-reflection methods build experience memories, but memories are mostly hard to invoke, refine, or scale, while agent skills offer a more actionable unit: structured procedural knowledge that specifies when to act, how to act, and which resources or tools to use. We introduce MASkills, a continual learning framework that optimizes multi-agent LLM systems through agent skills. MASkills presents a new agent-optimization pipeline that integrates skill-conditioned credit assignment, hierarchical credit aggregation, and momentum-smoothed optimization, enabling agent skill libraries to evolve through refinement, induction, consolidation, and pruning. Experiments on HotpotQA, LoCoMo, and GAIA demonstrate the effectiveness of MASkills across multiple agentic tasks. Our code is available at https://github.com/DaRL-GenAI/MASkills

Comment: 14 pages, 4 figures

Journal: EMNLP 2026 Findings

arXiv abs page · PDF · same-day batch