🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023Issue 180 · Sep 14 – Sep 20, 2026

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

2,315
Papers
180
Weekly issues
2023
Since
This week · 10 papersView the full issue →
AudioGPT

AudioGPT

Connects ChatGPT with audio foundational models for speech, music, sound, and talking head tasks.

02Multimodal
DataComp

DataComp

A multimodal dataset benchmark with 12.8B image-text pairs.

03Multimodal
ChatGPT for Information Extraction

ChatGPT for Information Extraction

A deeper assessment of ChatGPT on information extraction tasks.

04Evaluation
Comparing Physician vs ChatGPT (JAMA)

Comparing Physician vs ChatGPT (JAMA)

A JAMA Internal Medicine study comparing physician and ChatGPT responses.

05Evaluation
Stable and Low-Precision Training for Large-Scale Vision-Language Models

Stable and Low-Precision Training for Large-Scale Vision-Language Models

Methods for accelerating and stabilizing large VLM training.

06Training
DINOv2

DINOv2

Meta's self-supervised vision foundation model producing robust features without labels.

07Training
Learning to Compress Prompts with Gist Tokens

Learning to Compress Prompts with Gist Tokens

Trains LMs to compress prompts into reusable "gist" tokens.

08Efficiency
Scaling Biomolecular Simulations with Equivariant Models

Scaling Biomolecular Simulations with Equivariant Models

A framework for large-scale biomolecular simulation using equivariant deep learning.

09Architecture
Evaluating Verifiability in Generative Search Engines

Evaluating Verifiability in Generative Search Engines

Audits popular generative search engines for citation accuracy.

10Retrieval
Generative Disco: Text-to-Video Generation for Music Visualization

Generative Disco: Text-to-Video Generation for Music Visualization

An LLM + T2I system for music visualization.

11Multimodal
Architectures of Topological Deep Learning: A Survey on Topological Neural Networks

Architectures of Topological Deep Learning: A Survey on Topological Neural Networks

A comprehensive survey on topological neural networks.

12Architecture
Visual Instruction Tuning (LLaVA)

Visual Instruction Tuning (LLaVA)

Uses language-only GPT-4 to generate multimodal instruction-following data.

13Multimodal
ChatGPT: Applications, Opportunities, and Threats

ChatGPT: Applications, Opportunities, and Threats

A comprehensive overview of ChatGPT's applications and risks.

14Training
Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models

A framework inferring tool sequences for compositional reasoning.

15Reasoning
Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models

Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models

High-resolution video synthesis with latent diffusion.

16Multimodal
Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

Combines mip-NeRF 360 with grid-based models for 22x faster training.

17Training
Generative Agents: Interactive Simulacra of Human Behavior

Generative Agents: Interactive Simulacra of Human Behavior

Stanford/Google's landmark paper on LLM-powered social simulations.

18Agents
Emergent Autonomous Scientific Research Capabilities of LLMs

Emergent Autonomous Scientific Research Capabilities of LLMs

An agent combining LLMs for autonomous scientific experiments.

19Agents
Automatic Gradient Descent: Deep Learning without Hyperparameters

Automatic Gradient Descent: Deep Learning without Hyperparameters

A hyperparameter-free first-order optimizer that leverages architecture.

20Training
ChemCrow: Augmenting LLMs with Chemistry Tools

ChemCrow: Augmenting LLMs with Chemistry Tools

An LLM chemistry agent with 13 expert-designed tools.

21Agents
One Small Step for Generative AI, One Giant Leap for AGI

One Small Step for Generative AI, One Giant Leap for AGI

A complete survey on ChatGPT and GPT-4.

22Training
OpenAGI: When LLM Meets Domain Experts

OpenAGI: When LLM Meets Domain Experts

An open-source research platform for LLM agents manipulating domain expert models.

23Agents
AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

A benchmark using real human standardized exams.

24Evaluation
Teaching Large Language Models to Self-Debug

Teaching Large Language Models to Self-Debug

Teaches LLMs to debug their own code via few-shot demonstrations.

25Code
180 weeks of AI research · papers per week
Week of Sep 14–20, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Apr 2023Hover a week to inspect · select to openSep 2026