Ir al contenido principalSaltar al contenido

Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models

Abstract

arXiv:2608.05168v1 Announce Type: new Abstract: Large language models often fail on reasoning tasks despite possessing the capability to solve them. We argue that many such failures arise from localized reasoning bugs in intermediate steps rather than from global incompetence. We show that these bugs are frequently repairable: inserting a short patch generated by a weak probe model after the same strong-model reasoning prefix can redirect the trajectory toward a correct solution. However, this c

Más info: Política de Cookies