← 피드로
[Submitted on 11 Jun 2026]
Abstract:Chain-of-thought (CoT) reasoning is the dominant paradigm for inference-time scaling in language models, yet the causal influence of individual steps on the final answer poorly understood. We estimate each step’s causal importance via early exit and use this measure to study how answers form across the reasoning traces of several model families. Across diverse tasks, we find that reasoning typically crosses a emph{commitment boundary} — a sharp transition from transient intermediate guesses to a stable, high-confidence answer. This transition often happens in a single step, well before the model’s reasoning block ends, and is followed by emph{epiphenomenal} CoT steps that leave the final answer probability unaltered. Using attention probes, we show that answer-formation stages can be linearly decoded from intermediate reasoning steps with high accuracy and generalize robustly to unseen reasoning tasks. We exploit this signal to early-exit reasoning blocks at the commitment boundary, reducing the length of CoTs up to 55% on average with negligible impact on model performance.
Submission history
From: Daniel Scalena [view email]
[v1]
Thu, 11 Jun 2026 17:21:16 UTC (9,176 KB)
추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2606.13603
답글 남기기