[2509.23392] Your Models Have Thought Enough: Training Large Reasoning

[2509.23392] Your Models Have Thought Enough: Training Large Reasoning Models to Stop Overthinking

arXiv - AI March 31, 2026 4 min read

About this article

Abstract page for arXiv paper 2509.23392: Your Models Have Thought Enough: Training Large Reasoning Models to Stop Overthinking

Computer Science > Artificial Intelligence arXiv:2509.23392 (cs) [Submitted on 27 Sep 2025 (v1), last revised 30 Mar 2026 (this version, v3)] Title:Your Models Have Thought Enough: Training Large Reasoning Models to Stop Overthinking Authors:Jinyi Han, Ying Huang, Ying Liao, Zishang Jiang, Xikun Lu, Haiquan Zhao, Xinyi Wang, Guanghao Zhou, Sihang Jiang, Jiaqing Liang, Weikang Zhou, Zeye Sun, Fei Yu, Yanghua Xiao View a PDF of the paper titled Your Models Have Thought Enough: Training Large Reasoning Models to Stop Overthinking, by Jinyi Han and 13 other authors View PDF HTML (experimental) Abstract:Large Reasoning Models (LRMs) have achieved impressive performance on challenging tasks, yet their deep reasoning often incurs substantial computational costs. To achieve efficient reasoning, existing reinforcement learning methods still struggle to construct short reasoning path during the rollout stage, limiting effective learning. Inspired by Evidence Accumulation Models, we find that LRMs have accumulated sufficient information early in reasoning, making further reasoning steps redundant. Based on this insight, we propose Just-Enough Thinking (JET), which trains models to proactively terminate unnecessary reasoning. JET performs trajectory truncation during rollout to expose the model to short, distributionally consistent reasoning paths. Besides, it uses a quality-controlled length reward to better encourage concise reasoning while maintaining correctness. Extensive experim...

Originally published on March 31, 2026. Curated by AI News.

Ai Infrastructure

UMKC Announces New Master of Science in Artificial Intelligence

UMKC announces a new Master of Science in Artificial Intelligence program aimed at addressing workforce demand for AI expertise, set to l...

AI News - General · 4 min · about 1 hour ago

Machine Learning

AI assistants are optimized to seem helpful. That is not the same thing as being helpful.

RLHF trains models on human feedback. Humans rate responses they like. And it turns out humans consistently rate confident, fluent, agree...

Reddit - Artificial Intelligence · 1 min · about 1 hour ago

Llms

wtf bro did what? arc 3 2026

The Physarum Explorer is a high-speed, bio-inspired neural model designed specifically for ARC geometry. Here is the snapshot of its curr...

Reddit - Artificial Intelligence · 1 min · about 1 hour ago

Machine Learning

Meta Pauses Work With Mercor After Data Breach Puts AI Industry Secrets at Risk | WIRED

Major AI labs are investigating a security incident that impacted Mercor, a leading data vendor. The incident could have exposed key data...

Wired - AI · 6 min · about 1 hour ago

[2509.23392] Your Models Have Thought Enough: Training Large Reasoning Models to Stop Overthinking

About this article

Related Articles

UMKC Announces New Master of Science in Artificial Intelligence

AI assistants are optimized to seem helpful. That is not the same thing as being helpful.

wtf bro did what? arc 3 2026

Meta Pauses Work With Mercor After Data Breach Puts AI Industry Secrets at Risk | WIRED

No comments

Stay updated with AI News