🚀NEW LABGetting Started with Claude AgentsStart lab
Memory · Evaluation · Reasoning

Grok-1.5

Full-paper indexing in progress
Paper preview
Grok-1.5
Paper summary

xAI's Grok-1.5 is the successor to the open-weight Grok-1, emphasizing long-context understanding and substantially stronger math, code, and reasoning performance.

Ask this paper

Key points
01

Benchmarks: Reports 50.6% on the MATH benchmark, 90.0% on GSM8K, 74.1% on HumanEval, and 81.3% on MMLU - a large jump over Grok-1 and competitive with contemporary closed-source peers.

02

128K context window: Grok-1.5 can process up to 128K tokens, a 16x expansion over Grok-1, enabling longer documents and multi-turn sessions without external retrieval.

03

Long-context retrieval: On in-house needle-in-a-haystack evaluations the model demonstrates strong recall across its full 128K window, not just near the boundaries.

04

Availability: Rolled out first to early testers and existing Grok users on 𝕏, with broader release planned as the platform continues to iterate on reasoning and agentic features.

Every Monday
Get next week’s papers.
Subscribe on Substack