🚀NEW LABGetting Started with Claude AgentsStart lab
Reasoning

Contrastive Chain-of-Thought

First page
Contrastive Chain-of-Thought
Paper summary

Proposes contrastive CoT prompting where models see both valid *and* invalid reasoning demonstrations to reduce reasoning errors.

Ask this paper

Key points
01

Valid + invalid demos: Demonstrations pair correct reasoning traces with common incorrect ones, teaching the model what not to do as well as what to do.

02

Automatic construction: Provides an automatic method to generate contrastive demonstrations, avoiding the manual curation bottleneck that limited prior CoT variants.

03

Improves over CoT: Outperforms standard CoT across reasoning benchmarks, with particularly strong gains on problems where common error patterns are predictable.

04

Pedagogical analog: The improvement mirrors human learning research showing that studying worked examples and errors side-by-side beats studying successes alone.

Every Monday
Get next week’s papers.
Subscribe on Substack