NeFut Logo NeFut
Admin Login

[CS.AI] AgentBrew: Lifelong Knowledge Brewing from Strong Teachers to Weak LLM Agents

Published at: 2026-07-21 22:00 Last updated: 2026-07-22 01:01
#algorithm #AI #Machine Learning

Deploying large language model (LLM) agents typically requires a compact test-time student, even if a stronger teacher is available during training. We study knowledge brewing: distilling a teacher's interactive experience into a persistent external memory for the student. Crucially, this requires no weight updates, expert demonstrations, ground-truth labels, or test-time teacher access.

This setting poses two challenges: environments provide only sparse, binary feedback, and teacher-authored notes must be inherently tailored to be concretely executable by a substantially weaker student. To address these hurdles, we propose AgentBrew, comprising two coupled components.

First, a failure-triggered teacher—Ralph Loop mitigates sparse feedback by transforming student failures into environment-validated notes. Second, student-aware synthesis calibrates teacher knowledge to the weak executor's operational granularity, yielding model-specific, actionable guidance.

Extensive evaluations and comprehensive ablations across coding, math, and tool-use tasks demonstrate that this asymmetric, training-free brewing paradigm produces highly capable yet deployable LLM agents.

Blogger's Review: AgentBrew presents an innovative approach to address the limitations of LLM agents in real-world applications through knowledge transfer between teachers and students. The failure-triggered mechanism and targeted guidance design effectively enhance the execution capabilities of models in the absence of direct feedback, making it a noteworthy development.

Original Source: https://arxiv.org/abs/2607.16851

[h] Back to Home