🚀NEW LABGetting Started with Claude AgentsStart lab
Data

o1 Replication Journey

First page
o1 Replication Journey
Paper summary

reports to be replicating the capabilities of OpenAI's o1 model; their journey learning technique encourages learning not just shortcuts, but the complete exploration process, including trial and error, reflection, and backtracking; claims that with only 327 training samples, their journey learning technique surpassed shortcut learning by 8.0% on the MATH dataset.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack