🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
Agents · Reasoning

OPRO (LLMs as Optimizers)

Free while signed in. Answers cite the passages they came from.

First page
OPRO (LLMs as Optimizers)
The curator’s take

DeepMind's OPRO uses LLMs as general-purpose optimizers over natural-language-described problems.

Key points
01

Natural-language optimization: The optimization problem is described in natural language; the LLM iteratively proposes new solutions conditioned on previously found solutions.

02

Prompt optimization: As a key application, optimizes prompts to maximize test accuracy, using previously evaluated prompts as trajectory context.

03

Big gains over human prompts: LLM-optimized prompts outperform human-designed prompts on GSM8K and BIG-Bench Hard, sometimes by over 50 percentage points.

04

General-purpose pattern: Positions LLMs as general-purpose optimizers for problems that are hard to specify mathematically, including linear regression, traveling salesman variants, and prompt design.

Every Monday
Get next week’s papers.

The same picks and the same summaries, in your inbox. Free, and 176 issues deep.

Subscribe on Substack
OPRO (LLMs as Optimizers) | DAIR.AI Academy | DAIR.AI Academy