🚀NEW LABGetting Started with Claude AgentsStart lab
Retrieval

Self-RAG

First page
Self-RAG
Paper summary

Self-RAG trains an LM to adaptively retrieve, generate, and self-critique using special reflection tokens.

Ask this paper

Key points
01

Reflection tokens: Introduces special tokens that control retrieval decisions, passage relevance judgments, and self-evaluation of generations.

02

Adaptive retrieval: The model decides on-the-fly whether to retrieve, rather than always retrieving on every query - saving compute on knowledge-light queries.

03

Self-reflection: Critiques its own generations against retrieved passages, enabling controllable trade-offs between response quality and factuality at inference.

04

Significant gains: Outperforms state-of-the-art LLMs and strong RAG baselines on open-domain QA, reasoning, and fact verification.

Every Monday
Get next week’s papers.
Subscribe on Substack