🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
Training · Reasoning

Rethinking Reflection in Pre-Training

Free while signed in. Answers cite the passages they came from.

First page
Rethinking Reflection in Pre-Training
The curator’s take

Reflection — the ability of LLMs to identify and correct their own reasoning — has often been attributed to reinforcement learning or fine-tuning. This paper argues otherwise: reflection emerges during pre-training. The authors introduce adversarial reasoning tasks to show that self-reflection and correction capabilities steadily improve as compute increases, even in the absence of supervised post-training. Key contributions:

Key points
01

Propose two kinds of reflection:

02

Build six adversarial datasets (GSM8K, TriviaQA, CruxEval, BBH) to test reflection across math, coding, logic, and knowledge domains. On GSM8K-Platinum, explicit reflection rates grow from ~10% to 60% with increasing pre-training tokens.

03

Demonstrate that simple triggers like “Wait,” reliably induce reflection.

04

Evaluate 40 OLMo-2 and Qwen2.5 checkpoints, finding a strong correlation between pre-training compute and both accuracy and reflection rate. Why it matters:

05

Reflection is a precursor to reasoning and can develop before RLHF or test-time decoding strategies.

06

Implication: We can instill advanced reasoning traits with better pre-training data and scale, rather than relying entirely on post-training tricks.

07

They also show a trade-off: more training compute reduces the need for expensive test-time compute like long CoT traces.

Every Monday
Get next week’s papers.

The same picks and the same summaries, in your inbox. Free, and 176 issues deep.

Subscribe on Substack