跳到正文
原文
rohanpaul_ai· @rohanpaul_ai · X·· 2 天前AI 评分55

Meta、Duke 等机构论文提出分支化自改进 Agent harness 优化方法

AI 导读

Meta、Duke 和 California Univ 的新论文提出将 Agent harness 搜索拆分为专业化分支、按任务路由到最佳分支的方法,在全部 4 个测试设置中超过 Meta-Harness。

正文 · 原文

New paper from Meta, Duke, California Univ on self-Improving Agent's Harness Optimization

When an AI tunes your agent's harness, split the search into specialized branches and route each task to the best fit, which beat Meta-Harness in all 4 test settings.

Giving each tuning branch its own problems and its own notes on what worked produced harnesses with different strengths, and a router turned those strengths into higher scores.

A harness is the code around an LLM that controls its tools, retrieval and self-checks. Meta-Harness has an AI rewrite it in a loop, but every version is scored on the same problems, so search sticks to 1 path.

Each of 2 branches keeps the practice problems it solves better than the other and writes its own notes on what worked. In math, 1 branch learned to verify answers, while the other learned to build full derivations.

With Gemini 3 Flash on Olympiad math, accuracy rose from 46.0% with Meta-Harness to 62.0%.

If you auto-tune agents, keep several specialized harnesses and route between them rather than betting on 1 winner.

来源:rohanpaul_ai · x.com