🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

1,760
Papers
176
Weekly issues
2023
Since
390 papers · 2023Clear filters →
DINOv2

DINOv2

Meta's self-supervised vision foundation model producing robust features without labels.

361Training
Learning to Compress Prompts with Gist Tokens

Learning to Compress Prompts with Gist Tokens

Trains LMs to compress prompts into reusable "gist" tokens.

362Efficiency
Scaling Biomolecular Simulations with Equivariant Models

Scaling Biomolecular Simulations with Equivariant Models

A framework for large-scale biomolecular simulation using equivariant deep learning.

363Architecture
Evaluating Verifiability in Generative Search Engines

Evaluating Verifiability in Generative Search Engines

Audits popular generative search engines for citation accuracy.

364Retrieval
Generative Disco: Text-to-Video Generation for Music Visualization

Generative Disco: Text-to-Video Generation for Music Visualization

An LLM + T2I system for music visualization.

365Multimodal
Architectures of Topological Deep Learning: A Survey on Topological Neural Networks

Architectures of Topological Deep Learning: A Survey on Topological Neural Networks

A comprehensive survey on topological neural networks.

366Architecture
Visual Instruction Tuning (LLaVA)

Visual Instruction Tuning (LLaVA)

Uses language-only GPT-4 to generate multimodal instruction-following data.

367Multimodal
ChatGPT: Applications, Opportunities, and Threats

ChatGPT: Applications, Opportunities, and Threats

A comprehensive overview of ChatGPT's applications and risks.

368Training
Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

A framework inferring tool sequences for compositional reasoning.

369Reasoning
Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models

Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models

High-resolution video synthesis with latent diffusion.

370Multimodal
Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

Combines mip-NeRF 360 with grid-based models for 22x faster training.

371Training
Generative Agents: Interactive Simulacra of Human Behavior

Generative Agents: Interactive Simulacra of Human Behavior

Stanford/Google's landmark paper on LLM-powered social simulations.

372Agents
Emergent Autonomous Scientific Research Capabilities of LLMs

Emergent Autonomous Scientific Research Capabilities of LLMs

An agent combining LLMs for autonomous scientific experiments.

373Agents
Automatic Gradient Descent: Deep Learning without Hyperparameters

Automatic Gradient Descent: Deep Learning without Hyperparameters

A hyperparameter-free first-order optimizer that leverages architecture.

374Training
ChemCrow: Augmenting LLMs with Chemistry Tools

ChemCrow: Augmenting LLMs with Chemistry Tools

An LLM chemistry agent with 13 expert-designed tools.

375Agents
One Small Step for Generative AI, One Giant Leap for AGI

One Small Step for Generative AI, One Giant Leap for AGI

A complete survey on ChatGPT and GPT-4.

376Training
OpenAGI: When LLM Meets Domain Experts

OpenAGI: When LLM Meets Domain Experts

An open-source research platform for LLM agents manipulating domain expert models.

377Agents
AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

A benchmark using real human standardized exams.

378Evaluation
Teaching Large Language Models to Self-Debug

Teaching Large Language Models to Self-Debug

Teaches LLMs to debug their own code via few-shot demonstrations.

379Code
Segment Everything Everywhere All at Once (SEEM)

Segment Everything Everywhere All at Once (SEEM)

A promptable, interactive segmentation model.

380Multimodal
Segment Anything (SAM)

Segment Anything (SAM)

Meta's foundational model for image segmentation with massive training data release.

381Multimodal
Instruction Tuning with GPT-4

Instruction Tuning with GPT-4

Uses GPT-4 to generate instruction-following data for LLM fine-tuning.

382Training
Eight Things to Know about Large Language Models

Eight Things to Know about Large Language Models

Sam Bowman's influential primer on key LLM considerations.

383Evaluation
A Survey of Large Language Models

A Survey of Large Language Models

A 50-page comprehensive survey on LLMs.

384Training
176 weeks of AI research · papers per week
Week of Aug 17–23, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Apr 2023Hover a week to inspect · select to openAug 2026