跳到正文
原文
omarsar0· @omarsar0 · X·· 17 小时前AI 评分56

Harness-Aware Distillation 论文:让小模型 Agent 学会利用 harness 信息

AI 导读

DAIR.AI 的 Elvis Saravia 推荐一篇关于小语言模型 Agent 的 harness-aware distillation 论文(arxiv.org/abs/2610.02858)。

正文 · 原文

Really nice paper on harness-aware distillation for small language model agents.

(bookmark it)

It's a super interesting distillation framework for small language model agents that are deployed with a harness.

In simple terms, you run the big model with and without harness info, then train the small model on cases where its action changes.

Researchers show that adding the harness to on-policy distillation raises how often the student uses harness information on ALFWorld (65.7% to 73.1%) but leaves success flat (43.1% to 43.5%).

Their method, Harness-Aware Distillation, queries the same teacher with and without the harness information and trains the student to prefer the action chosen with it. A filter drops pairs whose preferred action contradicts the harness records. The method uses no task rewards or success labels.

The student reaches 63.4% on unseen ALFWorld tasks, compared with 47.0% for the best baseline, and exceeds its 8B teacher. It also escapes 59.7% of stalls, while the baselines stay near the untrained student's 46.8%.

Paper: arxiv.org/abs/2610.02858

Chat with Paper: academy.dair.ai/papers/harne…

来源:omarsar0 · x.com