🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

1,760
Papers
176
Weekly issues
2023
Since
390 papers · 2023Clear filters →
LLM360

LLM360

LLM360 is a framework for fully transparent open-source LLM development, with everything from data to training dynamics released.

25Training
LLMs in Medicine

LLMs in Medicine

A comprehensive survey (300+ papers) of LLMs applied to medicine, from clinical tasks to biomedical research.

26Evaluation
Beyond Human Data (ReST-EM)

Beyond Human Data (ReST-EM)

DeepMind's ReST-EM shows that model-generated data plus a reward function can substantially reduce dependence on human-generated data.

27Reasoning
Gaussian-SLAM

Gaussian-SLAM

A neural RGBD SLAM method that extends 3D Gaussian Splatting to achieve photorealistic scene reconstruction without sacrificing speed.

28Training
Pearl

Pearl

Meta's Pearl is a production-ready reinforcement learning agent package designed for real-world deployment constraints.

29Agents
QuIP#

QuIP#

Cornell's QuIP# is a 2-bit LLM quantization scheme that combines lattice codebooks with incoherence processing to close the quality gap to FP16.

30Efficiency
Gemini 1.0

Gemini 1.0

Google launches Gemini 1.0, a multimodal family natively designed to reason across text, images, video, audio, and code from the ground up.

31Multimodal
EfficientSAM

EfficientSAM

Meta's EfficientSAM is a lightweight Segment Anything variant that preserves most of SAM's zero-shot quality at a fraction of the compute.

32Training
Magicoder

Magicoder

Magicoder is a fully open-source code LLM that closes the gap with top commercial code models at only 7B parameters via high-quality synthetic instruction data.

33Code
LLMs on Graphs

LLMs on Graphs

A comprehensive overview of the many ways LLMs can be applied to graph-structured data and when each pattern is useful.

34Reasoning
Llama Guard

Llama Guard

Meta's Llama Guard is a compact, instruction-tuned safety classifier built on Llama 2-7B for input/output moderation in conversational AI.

35Safety
KTO (Kahneman-Tversky Optimization)

KTO (Kahneman-Tversky Optimization)

Contextual AI introduces KTO, an alignment objective derived from prospect theory that works with binary "good/bad" signals instead of preference pairs.

36Reinforcement Learning
Chain of Code

Chain of Code

DeepMind's Chain of Code extends CoT by encouraging LMs to write pseudocode that mixes real code with LM-simulated sub-routines.

37Reasoning
Data Management for LLMs

Data Management for LLMs

A survey of data-management research for LLM pretraining and supervised fine-tuning stages.

38Training
RankZephyr

RankZephyr

RankZephyr is an open-source LLM for listwise zero-shot reranking that bridges the effectiveness gap with GPT-4.

39Evaluation
The Efficiency Spectrum of LLMs

The Efficiency Spectrum of LLMs

A comprehensive review of algorithmic advancements for improving LLM efficiency across the full training-to-inference stack.

40Efficiency
GNoME

GNoME

DeepMind's Graph Networks for Materials Exploration (GNoME) is an AI system that discovered 2.2 million new crystal structures, including 380,000 thermodynamically stable ones.

41Agents
Open-Source LLMs vs. ChatGPT

Open-Source LLMs vs. ChatGPT

A survey cataloguing tasks where open-source LLMs claim to be on par with or better than ChatGPT.

42Evaluation
Adversarial Diffusion Distillation (SDXL Turbo)

Adversarial Diffusion Distillation (SDXL Turbo)

Stability AI's ADD trains a student diffusion model that produces high-quality images in just 1-4 sampling steps.

43Training
Seamless

Seamless

Meta's Seamless is a family of models for end-to-end expressive, streaming cross-lingual speech communication.

44Safety
MEDITRON-70B

MEDITRON-70B

EPFL's MEDITRON is an open-source family of medical LLMs at 7B and 70B parameters, continually pretrained on curated medical corpora.

45Training
Medprompt

Medprompt

Microsoft researchers show that careful prompt engineering can push general-purpose GPT-4 to state-of-the-art on medical benchmarks, no domain fine-tuning required.

46Evaluation
UniIR

UniIR

UniIR is a unified instruction-guided multimodal retriever that handles eight retrieval tasks across modalities with a single model.

47Multimodal
Safe Deployment of Generative AI (Nature)

Safe Deployment of Generative AI (Nature)

A Nature correspondence arguing that medical professionals - not commercial interests - must drive the development and deployment of generative AI in medicine.

48Safety
176 weeks of AI research · papers per week
Week of Aug 17–23, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Apr 2023Hover a week to inspect · select to openAug 2026