跳到正文
原文
X:Elvis Saravia (@omarsar0, DAIR.AI)· @omarsar0·· 1 天前AI 评分52

Microsoft 等提出 ScholarEvolve:用已发表研究演化 Agent harness

New paper from Microsoft and colleagues on evolving agent harnesses. It's a really cool idea to evolve a harness from published research. Something I have also been testing for the past couple of months. ScholarEvolve proposes harness changes based on published agent research rather than the agent's own failure logs. It splits the harness into modules for tool use, memory management, and task execution. It runs topic modeling over recent papers to identify distinct improvement strategies for each module, then implements and tests combinations. New papers can be added over time. With the model held fixed, Qwen3.5-27B goal completion on AppWorld Challenge rises from 49.6% to 63.6%, and GPT-5.4-mini on Tau2-Bench Telecom rises from 72.7% to 81.9%. Paper: https://arxiv.org/abs/2609.40169 Chat with Paper: https://academy.dair.ai/papers/learning-from-research-toward-lifelong-agent-harness-evolution-2609.40169

AI 导读

Microsoft 及合作者发布论文提出 ScholarEvolve,根据已发表的 Agent 研究而非 Agent 自身失败日志来提出 harness 改动。

来源:X:Elvis Saravia (@omarsar0, DAIR.AI) · x.lingyaoai.com