🚀NEW LABGetting Started with Claude AgentsStart lab
Data · Training · Code

Textbooks Are All You Need (phi-1)

First page
Textbooks Are All You Need (phi-1)
Paper summary

Introduces a 1.3B parameter code LLM trained on textbook-quality data.

Ask this paper

Key points
01

Data-quality thesis: Trained on a curated selection of textbook-quality web data plus synthetic textbooks/exercises generated with GPT-3.5.

02

Small model, strong HumanEval: Achieves 50.6% pass@1 on HumanEval despite being 1.3B - beating much larger models on code generation.

03

4-day training: Trained in just 4 days on 8 A100s, showing that aggressive data selection can substitute for massive compute.

04

Phi-series launch: Kicked off Microsoft's Phi-series (Phi-1.5, Phi-2, Phi-3) and catalyzed the "small-but-smart" model research program.

Every Monday
Get next week’s papers.
Subscribe on Substack