🚀NEW LABGetting Started with Claude AgentsStart lab
Safety · Evaluation

Claude 2

Full-paper indexing in progress
Paper preview
Claude 2
Paper summary

Anthropic's second-generation LLM with a detailed model card on safety, alignment, and capabilities.

Ask this paper

Key points
01

100K context: Launched with a 100K token context window, enabling document-scale reasoning use cases that were impractical with earlier models.

02

Safety evaluations: Comprehensive safety evaluations including harmlessness benchmarks, bias probes, and red-teaming results transparently disclosed.

03

Capabilities gains: Significant improvements on coding (71.2% HumanEval), math (GSM8k), and legal reasoning over Claude 1.3.

04

Consumer release: First Claude model available to consumers via claude.ai in the US and UK, broadening Anthropic's public footprint.

Every Monday
Get next week’s papers.
Subscribe on Substack