Rewarding efficient reasoning improves abstention on underspecified tasks in reasoning models
Read the original at arxiv.org→arXiv:2609.20846v1 Announce Type: new Abstract: While modern large reasoning models (LRMs) excel at providing correct answers in many tasks, we provide additional evidence for the observation that they often...
Original headline: "Rewarding Efficient Reasoning Improves Abstention on Underspecified Tasks in Reasoning Models"
Coverage timeline
- Sep 21, 04:00 UTC arXiv cs.CL lead source Rewarding Efficient Reasoning Improves Abstention on Underspecified Tasks in Reasoning Models