RESEARCH

Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models

ArXiv cs.AI · Fri, 07 Aug 2026 04:00:00 GMT

arXiv:2608.05168v1 Announce Type: new Abstract: Large language models often fail on reasoning tasks despite possessing the capability to solve them. We argue that many such failures arise from localized reasoning bugs in intermediate steps rather than from global incompetence. We

Read original source Discuss with SiiMON