🚀NEW LABGetting Started with Claude AgentsStart lab
Evaluation · Training · Data

Survey of LLMs

First page
Survey of LLMs
Paper summary

A survey that maps the landscape of the three dominant LLM families - GPT, Llama, and PaLM - and the shared toolbox used to build and augment them.

Ask this paper

Key points
01

Three-family framing: Organizes the field around GPT, Llama, and PaLM lineages, tracing how each family evolved since ChatGPT's November 2022 release.

02

Capabilities and techniques: Summarizes training techniques (pretraining, fine-tuning, RLHF) and augmentation methods (RAG, tool use, chain-of-thought) used across the three families.

03

Datasets and metrics: Catalogs the datasets used for training, fine-tuning, and evaluation, and compares popular evaluation benchmarks and metrics.

04

Open challenges: Closes with concrete research directions including efficiency, alignment, multilinguality, and evaluation that remain open for the next wave of LLM work.

Every Monday
Get next week’s papers.
Subscribe on Substack