🚀NEW LABGetting Started with Claude AgentsStart lab
DAIR.AI · Curated weekly since April 2023Issue 182 · Sep 28 – Oct 4, 2026

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

2,650
Papers
182
Weekly issues
2023
Since

Discover and explore top AI papers with Claude Code or Codex

npx @dair-ai/mcp setup
This week · 10 papersView the full issue →
DataComp

DataComp

A multimodal dataset benchmark with 12.8B image-text pairs.

02Multimodal
ChatGPT for Information Extraction

ChatGPT for Information Extraction

A deeper assessment of ChatGPT on information extraction tasks.

03Evaluation
Comparing Physician vs ChatGPT (JAMA)

Comparing Physician vs ChatGPT (JAMA)

A JAMA Internal Medicine study comparing physician and ChatGPT responses.

04Evaluation
Stable and Low-Precision Training for Large-Scale Vision-Language Models

Stable and Low-Precision Training for Large-Scale Vision-Language Models

Methods for accelerating and stabilizing large VLM training.

05Training
DINOv2

DINOv2

Meta's self-supervised vision foundation model producing robust features without labels.

06Training
Learning to Compress Prompts with Gist Tokens

Learning to Compress Prompts with Gist Tokens

Trains LMs to compress prompts into reusable "gist" tokens.

07Efficiency
Scaling Biomolecular Simulations with Equivariant Models

Scaling Biomolecular Simulations with Equivariant Models

A framework for large-scale biomolecular simulation using equivariant deep learning.

08Architecture
Evaluating Verifiability in Generative Search Engines

Evaluating Verifiability in Generative Search Engines

Audits popular generative search engines for citation accuracy.

09Retrieval
Generative Disco: Text-to-Video Generation for Music Visualization

Generative Disco: Text-to-Video Generation for Music Visualization

An LLM + T2I system for music visualization.

10Multimodal
Architectures of Topological Deep Learning: A Survey on Topological Neural Networks

Architectures of Topological Deep Learning: A Survey on Topological Neural Networks

A comprehensive survey on topological neural networks.

11Architecture
Visual Instruction Tuning (LLaVA)

Visual Instruction Tuning (LLaVA)

Uses language-only GPT-4 to generate multimodal instruction-following data.

12Multimodal
ChatGPT: Applications, Opportunities, and Threats

ChatGPT: Applications, Opportunities, and Threats

A comprehensive overview of ChatGPT's applications and risks.

13Training
Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

A framework inferring tool sequences for compositional reasoning.

14Reasoning
Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models

Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models

High-resolution video synthesis with latent diffusion.

15Multimodal
Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

Combines mip-NeRF 360 with grid-based models for 22x faster training.

16Training
Generative Agents: Interactive Simulacra of Human Behavior

Generative Agents: Interactive Simulacra of Human Behavior

Stanford/Google's landmark paper on LLM-powered social simulations.

17Agents
Emergent Autonomous Scientific Research Capabilities of LLMs

Emergent Autonomous Scientific Research Capabilities of LLMs

An agent combining LLMs for autonomous scientific experiments.

18Agents
Automatic Gradient Descent: Deep Learning without Hyperparameters

Automatic Gradient Descent: Deep Learning without Hyperparameters

A hyperparameter-free first-order optimizer that leverages architecture.

19Training
ChemCrow: Augmenting LLMs with Chemistry Tools

ChemCrow: Augmenting LLMs with Chemistry Tools

An LLM chemistry agent with 13 expert-designed tools.

20Agents
One Small Step for Generative AI, One Giant Leap for AGI

One Small Step for Generative AI, One Giant Leap for AGI

A complete survey on ChatGPT and GPT-4.

21Training
OpenAGI: When LLM Meets Domain Experts

OpenAGI: When LLM Meets Domain Experts

An open-source research platform for LLM agents manipulating domain expert models.

22Agents
AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

A benchmark using real human standardized exams.

23Evaluation
Teaching Large Language Models to Self-Debug

Teaching Large Language Models to Self-Debug

Teaches LLMs to debug their own code via few-shot demonstrations.

24Code
Segment Everything Everywhere All at Once (SEEM)

Segment Everything Everywhere All at Once (SEEM)

A promptable, interactive segmentation model.

25Multimodal
182 weeks of AI research · papers per week
Week of Sep 28–Oct 4, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Apr 2023Hover a week to inspect · select to openSep 2026