🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

1,761
Papers
176
Weekly issues
2023
Since
390 papers · 2023Clear filters →
TAPIR

TAPIR

Tracks any queried point on any physical surface throughout a video sequence faster than real-time.

289Evaluation
Mind2Web

Mind2Web

A dataset for evaluating generalist web agents with 2,350 tasks across 137 websites and 31 domains.

290Agents
Tracking Everything Everywhere All at Once (OmniMotion)

Tracking Everything Everywhere All at Once (OmniMotion)

Test-time optimization for dense, long-range motion estimation.

291Multimodal
AlphaDev

AlphaDev

DeepMind's deep RL agent discovering faster sorting algorithms from scratch, now in LLVM.

292Reinforcement Learning
Sparse-Quantized Representation (SpQR)

Sparse-Quantized Representation (SpQR)

Tim Dettmers' near-lossless LLM compression technique.

293Efficiency
MusicGen

MusicGen

A simple and controllable model for music generation using a single-stage Transformer.

294Multimodal
Augmenting LLMs with Databases (ChatDB)

Augmenting LLMs with Databases (ChatDB)

Combines an LLM with SQL databases as a symbolic memory framework.

295Memory
Concept Scrubbing in LLM (LEACE)

Concept Scrubbing in LLM (LEACE)

Least-squares Concept Erasure - erases a target concept from every layer of a neural network.

296Safety
Fine-Grained RLHF

Fine-Grained RLHF

Trains LMs with segment-level human feedback rather than whole-response preferences.

297Reinforcement Learning
Hierarchical Vision Transformer (Hiera)

Hierarchical Vision Transformer (Hiera)

Pretrains ViTs with MAE while removing unnecessary multi-stage complexity.

298Architecture
Humor in ChatGPT

Humor in ChatGPT

Explores ChatGPT's capabilities to grasp and reproduce humor.

299Evaluation
Imitating Reasoning Process of Larger LLMs (Orca)

Imitating Reasoning Process of Larger LLMs (Orca)

Microsoft's 13B model that imitates GPT-4's reasoning traces.

300Reasoning
Let's Verify Step by Step

Let's Verify Step by Step

OpenAI's landmark paper on process reward models for mathematical reasoning.

301Reasoning
No Positional Encodings (NoPE)

No Positional Encodings (NoPE)

Shows explicit position embeddings aren't essential for decoder-only Transformers.

302Architecture
BiomedGPT

BiomedGPT

A unified biomedical GPT for vision, language, and multimodal tasks.

303Multimodal
Thought Cloning

Thought Cloning

Imitation learning framework that learns to think as well as act.

304Agents
Fine-Tuning Language Models with Just Forward Passes (MeZO)

Fine-Tuning Language Models with Just Forward Passes (MeZO)

A memory-efficient zeroth-order optimizer for LLM fine-tuning.

305Training
MERT

MERT

An acoustic music understanding model with large-scale self-supervised training.

306Multimodal
Bytes Are All You Need

Bytes Are All You Need

Performs classification directly on file bytes without decoding.

307Training
Direct Preference Optimization (DPO)

Direct Preference Optimization (DPO)

Rafailov et al.'s simpler alternative to RLHF that rivals full RL-based alignment.

308Reinforcement Learning
SQL-PaLM

SQL-PaLM

An LLM-based Text-to-SQL system built on PaLM-2.

309Code
CodeTF

CodeTF

An open-source Transformer library for state-of-the-art code LLMs.

310Code
QLoRA

QLoRA

Tim Dettmers' breakthrough technique enabling 65B LLM fine-tuning on a single 48GB GPU.

311Training
LIMA

LIMA

Meta's 65B LLaMA fine-tuned on just 1,000 curated examples - showing alignment needs less data than believed.

312Training
176 weeks of AI research · papers per week
Week of Aug 17–23, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Apr 2023Hover a week to inspect · select to openAug 2026