🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

1,760
Papers
176
Weekly issues
2023
Since
618 papers · TrainingClear filters →
Synthetic Data Reduces Sycophancy

Synthetic Data Reduces Sycophancy

Google shows that fine-tuning on simple synthetic data can significantly reduce LLM sycophancy.

553Data
Open Problems and Limitations of RLHF

Open Problems and Limitations of RLHF

A comprehensive survey of open problems and fundamental limitations of RLHF as an alignment approach.

554Reinforcement Learning
Med-Flamingo

Med-Flamingo

Stanford's Med-Flamingo is a multimodal medical model supporting in-context learning for few-shot medical visual QA.

555Multimodal
AutoRobotics-Zero

AutoRobotics-Zero

Discovers zero-shot adaptable robot policies from scratch, including the automatic discovery of Python control code.

556Robotics
RT-2

RT-2

Google DeepMind's end-to-end vision-language-action model that learns from both web and robotics data to control robots.

557Robotics
Tracking Anything in High Quality

Tracking Anything in High Quality

A framework for high-quality tracking-anything in videos combining segmentation and refinement.

558Multimodal
Foundation Models in Vision

Foundation Models in Vision

A comprehensive survey on foundational models for computer vision and their open research directions.

559Multimodal
LoraHub

LoraHub

Enables efficient cross-task generalization via dynamic LoRA composition.

560Training
Survey of Aligned LLMs

Survey of Aligned LLMs

A comprehensive overview of alignment approaches covering data, training, and evaluation.

561Safety
Llama 2

Llama 2

Meta's open-weight foundation model family with chat-tuned variants ranging from 7B to 70B parameters.

562Training
CM3Leon

CM3Leon

Meta's retrieval-augmented multi-modal language model that generates both text and images.

563Multimodal
Secrets of RLHF in LLMs

Secrets of RLHF in LLMs

A deep investigation into RLHF with a focus on the inner workings of PPO, including open-source code.

564Reinforcement Learning
AnimateDiff

AnimateDiff

Animates frozen text-to-image diffusion models via a plug-in motion modeling module.

565Multimodal
Generative Pretraining in Multimodality (Emu)

Generative Pretraining in Multimodality (Emu)

A transformer-based multimodal foundation model for generating images and text.

566Multimodal
Multimodal Generation with Frozen LLMs

Multimodal Generation with Frozen LLMs

Maps images to LLM token space enabling models like PaLM and GPT-4 to handle visual tasks without parameter updates.

567Multimodal
CodeGen2.5

CodeGen2.5

Salesforce's new 7B code LLM trained on 1.5T tokens and optimized for fast sampling.

568Code
Extending Context Window of LLMs (PI)

Extending Context Window of LLMs (PI)

Position Interpolation extends LLaMA's context to 32K with minimal fine-tuning (within 1000 steps).

569Memory
Visual Navigation Transformer (ViNT)

Visual Navigation Transformer (ViNT)

A foundation model for vision-based robotic navigation built on flexible Transformers.

570Robotics
Textbooks Are All You Need (phi-1)

Textbooks Are All You Need (phi-1)

Introduces a 1.3B parameter code LLM trained on textbook-quality data.

571Data
RoboCat

RoboCat

DeepMind's self-improving foundation agent that operates different robotic arms from as few as 100 demonstrations.

572Robotics
ClinicalGPT

ClinicalGPT

A language model optimized through extensive and diverse medical data and multi-turn dialogue.

573Training
LOMO

LOMO

A memory-efficient optimizer that combines gradient computation and parameter update in one step.

574Efficiency
LMFlow

LMFlow

An extensible and lightweight toolkit for fine-tuning and inference of large foundation models.

575Training
FinGPT

FinGPT

An open-source LLM for the finance sector with a data-centric approach.

576Evaluation
176 weeks of AI research · papers per week
Week of Aug 17–23, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Apr 2023Hover a week to inspect · select to openAug 2026