🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

1,760
Papers
176
Weekly issues
2023
Since
390 papers · 2023Clear filters →
Active Retrieval Augmented LLMs (FLARE)

Active Retrieval Augmented LLMs (FLARE)

Actively decides when and what to retrieve during generation.

337Retrieval
FrugalGPT

FrugalGPT

Strategies to reduce LLM inference cost while improving performance.

338Efficiency
StarCoder

StarCoder

An open-access 15.5B code LLM with 8K context and 80+ programming languages.

339Code
MultiModal-GPT

MultiModal-GPT

A vision-language model for multi-round dialogue fine-tuned from OpenFlamingo.

340Multimodal
scGPT

scGPT

A foundation model for single-cell multi-omics pretrained on 10 million cells.

341Training
GPTutor

GPTutor

A ChatGPT-powered VSCode extension for code explanation.

342Training
Shap-E

Shap-E

OpenAI's conditional generative model for 3D assets producing implicit functions.

343Multimodal
Are Emergent Abilities of LLMs a Mirage?

Are Emergent Abilities of LLMs a Mirage?

Stanford's critical re-examination of emergent abilities.

344Evaluation
Interpretable ML for Science with PySR

Interpretable ML for Science with PySR

An open-source library for practical symbolic regression in the sciences.

345Safety
PMC-LLaMA

PMC-LLaMA

A LLaMA model fine-tuned on 4.8 million medical papers.

346Training
Distilling Step-by-Step!

Distilling Step-by-Step!

A mechanism to train smaller models that outperform larger LLMs using fewer examples.

347Training
Poisoning Language Models During Instruction Tuning

Poisoning Language Models During Instruction Tuning

Shows adversaries can poison LLMs via instruction tuning data.

348Training
Unlimiformer

Unlimiformer

Long-range Transformers with unlimited length input via external datastores.

349Retrieval
Learning to Reason and Memorize with Self-Notes

Learning to Reason and Memorize with Self-Notes

LLMs that deviate from input to explicitly "think" and memorize.

350Reasoning
Learning Agile Soccer Skills for a Bipedal Robot with Deep RL

Learning Agile Soccer Skills for a Bipedal Robot with Deep RL

DeepMind's bipedal humanoid robot playing soccer.

351Robotics
Scaling Transformer to 1M tokens with RMT

Scaling Transformer to 1M tokens with RMT

Recurrent Memory Transformer extends BERT's effective context to 2M tokens.

352Memory
Track Anything

Track Anything

An interactive tool for video object tracking and segmentation built on Segment Anything.

353Multimodal
A Cookbook of Self-Supervised Learning

A Cookbook of Self-Supervised Learning

A comprehensive overview of SSL techniques and practical considerations.

354Training
Harnessing the Power of LLMs in Practice: A Survey on ChatGPT and Beyond

Harnessing the Power of LLMs in Practice: A Survey on ChatGPT and Beyond

A practical guide for practitioners working with LLMs.

355Evaluation
AudioGPT

AudioGPT

Connects ChatGPT with audio foundational models for speech, music, sound, and talking head tasks.

356Multimodal
DataComp

DataComp

A multimodal dataset benchmark with 12.8B image-text pairs.

357Multimodal
ChatGPT for Information Extraction

ChatGPT for Information Extraction

A deeper assessment of ChatGPT on information extraction tasks.

358Evaluation
Comparing Physician vs ChatGPT (JAMA)

Comparing Physician vs ChatGPT (JAMA)

A JAMA Internal Medicine study comparing physician and ChatGPT responses.

359Evaluation
Stable and Low-Precision Training for Large-Scale Vision-Language Models

Stable and Low-Precision Training for Large-Scale Vision-Language Models

Methods for accelerating and stabilizing large VLM training.

360Efficiency
176 weeks of AI research · papers per week
Week of Aug 17–23, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Apr 2023Hover a week to inspect · select to openAug 2026