AdviSD: Advisor model learns to steer frozen LLMs via multi‑turn self‑distillation
A small advisor model is trained to give natural‑language advice to a frozen language model executor, improving performance through self‑distillation.
A small advisor model is trained to give natural‑language advice to a frozen language model executor, improving performance through self‑distillation.