Harness-Aware Distillation trains small model agents to 63.4% success on ALFWorld
A research paper presents Harness-Aware Distillation, a training framework that teaches small language model agents to leverage harness information by contrasting teacher decisions. The resulting student model reached 63.4% on unseen ALFWorld tasks, outperforming its 8B teacher.