🚀NEW LABGetting Started with Claude AgentsStart lab
Reasoning · Evaluation

Mathematical LLMs Survey

First page
Mathematical LLMs Survey
Paper summary

A survey on the progress of LLMs on mathematical reasoning tasks, covering methods, benchmarks, and open problems.

Ask this paper

Key points
01

Task taxonomy: Covers math word problem solving, symbolic reasoning, and theorem proving, showing which capabilities emerge at which model scales.

02

Methods landscape: Reviews prompting techniques (CoT, PoT, ToT, self-verification) alongside fine-tuning and tool-use approaches.

03

Dataset reference: Catalogs the dominant math benchmarks (GSM8K, MATH, MiniF2F, etc.) and their evaluation methodologies.

04

Frontier problems: Highlights reasoning-faithfulness, formal-vs-informal math integration, and reward-model design as the key open questions.

Every Monday
Get next week’s papers.
Subscribe on Substack