🚀NEW LABGetting Started with Claude AgentsStart lab
Reasoning · Evaluation

PrompCoT 2.0

First page
PrompCoT 2.0
Paper summary

PromptCoT 2.0 introduces an EM-based loop for synthesizing harder and more diverse reasoning prompts, replacing manual heuristics from PromptCoT 1.0. It enables both self-play and SFT training regimes, achieving new SOTA on reasoning benchmarks like AIME, HMMT, LiveCodeBench, and Codeforces, showing prompt synthesis as a new scaling axis for LLM reasoning.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack