Paper Detail

MASkills: Continual Skills Optimization for Multi-Agent LLM Systems

Huaiyuan Yao, Xiaoou Liu, Charles Fleming, Tianlong Chen, Hua Wei

arxiv Score 18.5

Published 2026-09-02 · First seen 2026-09-03

Research Track A · General AI

Abstract

LLM-based multi-agent systems have shown strong performance on complex tasks, yet continual improvement from interaction experience remains challenging. Existing self-reflection methods build experience memories, but memories are mostly hard to invoke, refine, or scale, while agent skills offer a more actionable unit: structured procedural knowledge that specifies when to act, how to act, and which resources or tools to use. We introduce MASkills, a continual learning framework that optimizes multi-agent LLM systems through agent skills. MASkills presents a new agent-optimization pipeline that integrates skill-conditioned credit assignment, hierarchical credit aggregation, and momentum-smoothed optimization, enabling agent skill libraries to evolve through refinement, induction, consolidation, and pruning. Experiments on HotpotQA, LoCoMo, and GAIA demonstrate the effectiveness of MASkills across multiple agentic tasks. Our code is available at https://github.com/DaRL-GenAI/MASkills

Workflow Status

Review status
pending
Role
unreviewed
Read priority
now
Vote
Not set.
Saved
no
Collections
Not filed yet.
Next action
Not filled yet.

Reading Brief

No structured notes yet. Add `summary_sections`, `why_relevant`, `claim_impact`, or `next_action` in `papers.jsonl` to enrich this view.

Why It Surfaced

No ranking explanation is available yet.

Tags

No tags.

BibTeX

@article{yao2026maskills,
  title = {MASkills: Continual Skills Optimization for Multi-Agent LLM Systems},
  author = {Huaiyuan Yao and Xiaoou Liu and Charles Fleming and Tianlong Chen and Hua Wei},
  year = {2026},
  abstract = {LLM-based multi-agent systems have shown strong performance on complex tasks, yet continual improvement from interaction experience remains challenging. Existing self-reflection methods build experience memories, but memories are mostly hard to invoke, refine, or scale, while agent skills offer a more actionable unit: structured procedural knowledge that specifies when to act, how to act, and which resources or tools to use. We introduce MASkills, a continual learning framework that optimizes mu},
  url = {https://arxiv.org/abs/2609.02094},
  keywords = {cs.AI, cs.CL},
  eprint = {2609.02094},
  archiveprefix = {arXiv},
}

Metadata

{}