EvoAgent: An Evolvable Agent Framework with Skill Learning and Multi-Agent Delegation

arXiv cs.AI / 4/23/2026

📰 NewsIdeas & Deep AnalysisModels & Research

共有:

Key Points

EvoAgentは、構造化されたスキル学習と、階層的なサブエージェント委譲を統合した「進化可能な」LLMエージェントのフレームワークです。
スキルはマルチファイルの能力ユニットとして表現され、トリガー機構や進化に関するメタデータを備えたうえで、ユーザーフィードバックに基づくクローズドループで継続的に生成・最適化されます。
3段階のスキル照合戦略と3層メモリにより、複雑課題を動的に分解し、長期的な能力の蓄積を支えます。
実世界の貿易関連シナリオの実験では、GPT5.2にEvoAgentを組み込むことで専門性・正確性・実用性が大きく向上し、LLM-as-Judge評価で平均スコアが約28%上昇したと報告されています。
モデル転移の追加実験では、エージェント性能は基盤モデルの能力だけでなく、モデルとエージェント設計（アーキテクチャ）の「シナジー」の大きさにも左右されることが示唆されています。

Abstract

This paper proposes EvoAgent - an evolvable large language model (LLM) agent framework that integrates structured skill learning with a hierarchical sub-agent delegation mechanism. EvoAgent models skills as multi-file structured capability units equipped with triggering mechanisms and evolutionary metadata, and enables continuous skill generation and optimization through a user-feedback-driven closed-loop process. In addition, by incorporating a three-stage skill matching strategy and a three-layer memory architecture, the framework supports dynamic task decomposition for complex problems and long-term capability accumulation. Experimental results based on real-world foreign trade scenarios demonstrate that, after integrating EvoAgent, GPT5.2 achieves significant improvements in professionalism, accuracy, and practical utility. Under a five-dimensional LLM-as-Judge evaluation protocol, the overall average score increases by approximately 28%. Further model transfer experiments indicate that the performance of an agent system depends not only on the intrinsic capabilities of the underlying model, but also on the degree of synergy between the model and the agent architecture.

I’m working on an AGI and human council system that could make the world better and keep checks and balances in place to prevent catastrophes. It could change the world. Really. Im trying to get ahead of the game before an AGI is developed by someone who only has their best interest in mind.

Reddit r/artificial

Deepseek V4 Flash and Non-Flash Out on HuggingFace

Reddit r/LocalLLaMA

DeepSeek V4 Flash & Pro Now out on API

Reddit r/LocalLLaMA

I’m building a post-SaaS app catalog on Base, and here’s what that actually means

Dev.to

From "Hello World" to "Hello Agents": The Developer Keynote That Rewired Software Engineering

Dev.to

EvoAgent: An Evolvable Agent Framework with Skill Learning and Multi-Agent Delegation

Key Points

Abstract

Related Articles

I’m working on an AGI and human council system that could make the world better and keep checks and balances in place to prevent catastrophes. It could change the world. Really. Im trying to get ahead of the game before an AGI is developed by someone who only has their best interest in mind.

Deepseek V4 Flash and Non-Flash Out on HuggingFace

DeepSeek V4 Flash & Pro Now out on API

I’m building a post-SaaS app catalog on Base, and here’s what that actually means

From "Hello World" to "Hello Agents": The Developer Keynote That Rewired Software Engineering

関連おすすめサービス

Notta搭載AI議事録イヤホン ZENCHORD1

AI搭載ボイスレコーダー Plaud

画像高画質化AIツール Aiarty Image Enhancer