🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

1,760
Papers
176
Weekly issues
2023
Since
618 papers · TrainingClear filters →
PMC-LLaMA

PMC-LLaMA

A LLaMA model fine-tuned on 4.8 million medical papers.

601Training
Distilling Step-by-Step!

Distilling Step-by-Step!

A mechanism to train smaller models that outperform larger LLMs using fewer examples.

602Training
Poisoning Language Models During Instruction Tuning

Poisoning Language Models During Instruction Tuning

Shows adversaries can poison LLMs via instruction tuning data.

603Training
Unlimiformer

Unlimiformer

Long-range Transformers with unlimited length input via external datastores.

604Retrieval
A Cookbook of Self-Supervised Learning

A Cookbook of Self-Supervised Learning

A comprehensive overview of SSL techniques and practical considerations.

605Training
ChatGPT for Information Extraction

ChatGPT for Information Extraction

A deeper assessment of ChatGPT on information extraction tasks.

606Evaluation
DINOv2

DINOv2

Meta's self-supervised vision foundation model producing robust features without labels.

607Training
Visual Instruction Tuning (LLaVA)

Visual Instruction Tuning (LLaVA)

Uses language-only GPT-4 to generate multimodal instruction-following data.

608Multimodal
ChatGPT: Applications, Opportunities, and Threats

ChatGPT: Applications, Opportunities, and Threats

A comprehensive overview of ChatGPT's applications and risks.

609Training
Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance Fields

Combines mip-NeRF 360 with grid-based models for 22x faster training.

610Training
Automatic Gradient Descent: Deep Learning without Hyperparameters

Automatic Gradient Descent: Deep Learning without Hyperparameters

A hyperparameter-free first-order optimizer that leverages architecture.

611Training
One Small Step for Generative AI, One Giant Leap for AGI

One Small Step for Generative AI, One Giant Leap for AGI

A complete survey on ChatGPT and GPT-4.

612Training
Instruction Tuning with GPT-4

Instruction Tuning with GPT-4

Uses GPT-4 to generate instruction-following data for LLM fine-tuning.

613Training
A Survey of Large Language Models

A Survey of Large Language Models

A 50-page comprehensive survey on LLMs.

614Training
Baize: An Open-Source Chat Model with Self-Chat Data

Baize: An Open-Source Chat Model with Self-Chat Data

An open chat model fine-tuned with LoRA on self-chat dialogs.

615Training
Better Language Models of Code through Self-Improvement

Better Language Models of Code through Self-Improvement

Self-improving code LLMs via pseudo-data generation.

616Code
Summary of ChatGPT/GPT-4 Research

Summary of ChatGPT/GPT-4 Research

An overview of ChatGPT and GPT-4 applications based on 194 papers.

617Training
Pythia

Pythia

EleutherAI's suite for analyzing LLMs across training and scaling.

618Training
176 weeks of AI research · papers per week
Week of Aug 17–23, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Apr 2023Hover a week to inspect · select to openAug 2026