NeFut Logo NeFut
中 Admin Login

[CS.AI] Harness-Aware Distillation for Small Language Model Agents

Published at: 2026-10-05 22:00 Last updated: 2026-10-06 12:11
#AI #Machine Learning #LLM

Language model agents are usually wrapped with a harness that manages context, tools, and feedback. When a large‑scale teacher agent is distilled into a smaller student, the harness stays unchanged, so the student only needs to acquire abilities that go beyond what the harness provides, such as correctly interpreting harness information. Standard distillation simply imitates the teacher’s full output and treats the harness as ordinary input, which prevents the student from learning the teacher’s extra contribution. To address this, we introduce Harness‑Aware Distillation (HAD), which isolates the teacher’s value beyond the harness. HAD augments on‑policy distillation with two components:

The contrast supplies information that pure imitation cannot provide, and HAD requires no task rewards, success labels, or future information. Experiments on several long‑horizon agent benchmarks and across different model sizes show that, with the same fixed harness, HAD consistently outperforms on‑policy distillation baselines. Further analysis reveals that HAD enters fewer unproductive loops and recovers from errors more often than baselines, suggesting it retains learnable feedback in its weights while still reading state information from the harness.

Review

Original Source: https://arxiv.org/abs/2610.02858

[h] Back to Home