Rewarding Efficient Reasoning Improves Abstention on Underspecified Tasks in Reasoning Models
arXiv:2609.20846v1 Announce Type: new Abstract: While modern large reasoning models (LRMs) excel at providing correct answers in many tasks, we provide additional evidence for the observation that they often struggle with a critical capability: knowing when to abstain from answering. We analyze thi…
Read the full story at arXiv cs.CL ↗
Timeline · 1 report
- 2026-09-21 04:00 · arXiv cs.CL
Rewarding Efficient Reasoning Improves Abstention on Underspecified Tasks in Reasoning Models