🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

1,760
Papers
176
Weekly issues
2023
Since
614 papers · ReasoningClear filters →
The Larger They Are, the Harder They Fail

The Larger They Are, the Harder They Fail

Reveals inverse-scaling failures in LLM code generation.

601Code
LLM Research Directions

LLM Research Directions

A list of research directions for students entering LLM research.

602Evaluation
Evidence of Meaning in Language Models Trained on Programs

Evidence of Meaning in Language Models Trained on Programs

Argues LMs learn meaning despite only next-token prediction.

603Reasoning
Towards Expert-Level Medical Question Answering (Med-PaLM 2)

Towards Expert-Level Medical Question Answering (Med-PaLM 2)

Google's second-generation medical LLM.

604Evaluation
StructGPT

StructGPT

A general framework for LLM reasoning over structured data.

605Reasoning
TinyStories

TinyStories

Explores how small LMs can be and still speak coherent English.

606Data
Symbol tuning

Symbol tuning

Fine-tunes LMs on in-context input-label pairs with natural-language labels replaced by arbitrary symbols.

607Reasoning
PaLM 2

PaLM 2

Google's second-generation PaLM powering Bard and Google products.

608Reasoning
Unfaithful Explanations in Chain-of-Thought Prompting

Unfaithful Explanations in Chain-of-Thought Prompting

Demonstrates CoT explanations can misrepresent the true reason for a model's prediction.

609Reasoning
StarCoder

StarCoder

An open-access 15.5B code LLM with 8K context and 80+ programming languages.

610Code
Distilling Step-by-Step!

Distilling Step-by-Step!

A mechanism to train smaller models that outperform larger LLMs using fewer examples.

611Training
Learning to Reason and Memorize with Self-Notes

Learning to Reason and Memorize with Self-Notes

LLMs that deviate from input to explicitly "think" and memorize.

612Reasoning
Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

A framework inferring tool sequences for compositional reasoning.

613Reasoning
Teaching Large Language Models to Self-Debug

Teaching Large Language Models to Self-Debug

Teaches LLMs to debug their own code via few-shot demonstrations.

614Code
176 weeks of AI research · papers per week
Week of Aug 17–23, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Apr 2023Hover a week to inspect · select to openAug 2026