Woodpecker Distillation shows weak models diagnose reasoning bugs in strong models
Read the original at arxiv.org→arXiv:2608.05168v1 Announce Type: new Abstract: Large language models often fail on reasoning tasks despite possessing the capability to solve them. We argue that many such failures arise from localized reasoning...
Original headline: "Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models"
Coverage timeline
- Aug 7, 04:00 UTC arXiv cs.AI lead source Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models