Skip to content
TrackPodcasts
scienceOct 1, 202524:47pending

When Does Reasoning Matter? A Controlled Study of Reasoning's Contribution to Model Performance

About this episode

🤗 Upvotes: 29 | cs.CL

Authors:
Nicolas Boizard, Hippolyte Gisserot-Boukhlef, Kevin El-Haddad, Céline Hudelot, Pierre Colombo

Title:
When Does Reasoning Matter? A Controlled Study of Reasoning's Contribution to Model Performance

Arxiv:
http://arxiv.org/abs/2509.22193v1

Abstract:
Large Language Models (LLMs) with reasoning capabilities have achieved state-of-the-art performance on a wide range of tasks. Despite its empirical success, the tasks and model scales at which reasoning becomes effective, as well as its training and inference costs, remain underexplored. In this work, we rely on a synthetic data distillation framework to conduct a large-scale supervised study. We compare Instruction Fine-Tuning (IFT) and reasoning models of varying sizes, on a wide range of math-centric and general-purpose tasks, evaluating both multiple-choice and open-ended formats. Our analysis reveals that reasoning consistently improves model performance, often matching or surpassing significantly larger IFT systems. Notably, while IFT remains Pareto-optimal in training and inference costs, reasoning models become increasingly valuable as model size scales, overcoming IFT performance limits on reasoning-intensive and open-ended tasks.

Get every episode summarized

Each time Daily Paper Cast publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

When Does Reasoning Matter? A Controlled Study of Reasoning's Contribution to Model Performance

Daily Paper Cast

0:00
24:47

More episodes

More from Daily Paper Cast

View all episodes →