Skip to content

GraphSkillEvo: Evolutionary Optimization of Graph-Structured Agent Skills

Sep 2026 · 0 citations · 91 references
Computer Science

TL;DR

GraphSkillEvo is introduced, a population-based evolutionary optimization framework with mutation and crossover operators for graph-structured skills that enables broader and more comprehensive exploration of the structured skill space than purely LLM-based iterative self-refinement.

Abstract

Skills can improve the performance of Large Language Model (LLM) agents by providing task-specific procedural guidance, while skill optimization further improves their effectiveness through iterative refinement. However, existing skill optimization methods typically represent skills as unstructured natural-language instructions, creating two key challenges: 1) Unstructured skills often lack explicit workflow-level guidance and contain substantial redundancy, making them difficult for LLMs to execute; 2) the vast search space of unconstrained natural-language skills makes skill optimization ineffective. To address these challenges, we propose representing skills as graph-structured natural-language artifacts. In graph-structured skills, each node represents an execution step together with its operational guidance, while directed edges encode context-dependent transitions between steps. Compared to unstructured skills, graph-structured skills can provide clear workflow-level guidance. Moreover, the proposed graph-structured skill can also facilitate skill optimization. Building on this structured representation, we introduce GraphSkillEvo, a population-based evolutionary optimization framework with mutation and crossover operators for graph-structured skills. By maintaining multiple candidate skills and combining effective components, GraphSkillEvo enables broader and more comprehensive exploration of the structured skill space than purely LLM-based iterative self-refinement. Extensive experiments across five agent benchmarks demonstrate that GraphSkillEvo consistently outperforms the strong skill optimization baseline SkillOpt, improving average accuracy by 4.01% on GPT-5.4-nano and 1.76% on GPT-5.4. Our code is available at https://github.com/ruisun7/GraphSkillEvo.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

GTA: Graph Theory Agent and Benchmark for Algorithmic Graph Reasoning with LLMs

The Graph Theory Agent (GTA), which pairs a preference-trained representation selector with plan-and-decompose scaffolding around a frozen executor LLM, is proposed, which lifts Phi-4 from 53.5% to 69.1% on the benchmark's easy split and from 33.0% to 41.5% on its hard split.

Zi-Xiang Xu, Yan-Bo Wang, Chenxi Wang et al. · 2 citations · ⚡1
#artificial intelligence Preprint Sep 2026

Rep2Skill: Representation-Guided Skill Self-Evolution for LLM Agents

Textual skills enable large language model (LLM) based agents to accumulate reusable procedural knowledge without updating model parameters. Yet existing skill evolution remains largely confined to the text space: an optimizer must diagnose success and failure patterns, and revise skills solely from long execution traj...

Kai-Xin Zhang, Chang-Ming Li, Ying-Dong Shi et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

The Procedural Graph is introduced: just as a knowledge graph organizes factual knowledge into (entity, relation, entity) triplets for what-is questions, a Procedural Graph organizes procedural knowledge into (procedure, relation, procedure) triplets for what-to-do questions.

Yu-Xing Lu, Yi-Cheng Chen, Shan-Chan Wu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Inference-Time Graph Engineering for Multi-Agent LLM Workflows

This work synthesizes a task-conditioned temporal workflow graph that jointly specifies agent connectivity and edge-level communication semantics, and introduces ReActNet, a training-free framework that compiles a query and a set of role-specialized agents into a sequence of directed communication graphs.

Katherine Tieu, Dong-Qi Fu, Ying-Long Xia et al. · 1 citation · ⚡1
#artificial intelligence Preprint Sep 2026

SkillVine: Agent Skill Evolution via Branching Exploration

SkillVine is proposed, an automatic skill-evolution framework that formulates skill evolution as a graph search problem and employs a branching exploration strategy, and achieves a balance between exploration and exploitation.

Kai-Wei Liu, Ji-Qian Dong, Li-Ran Dong et al. · 0 citations
#machine learning Preprint Sep 2026

SkillSpec: Consensus-Gated Agent Skill Evolution via Representation Specialization

Natural-language skills are textual procedural memories through which large language model (LLM) agents retain reusable task knowledge without updating model weights. Existing methods typically treat skills as either static artifacts or monolithic documents optimized using aggregate validation scores as feedback. Howev...

Huan-Cheng Chen, Xiao-Di Sun, Zhao-Qiong Huang et al. · 0 citations

Related blog posts

MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.