🚀NEW LABGetting Started with Claude AgentsStart lab
DAIR.AI · Curated weekly since April 2023Issue 182 · Sep 28 – Oct 4, 2026

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

2,650
Papers
182
Weekly issues
2023
Since

Discover and explore top AI papers with Claude Code or Codex

npx @dair-ai/mcp setup
This week · 10 papersView the full issue →
Retentive Network (RetNet)

Retentive Network (RetNet)

Microsoft's proposed foundation architecture aiming to replace Transformer attention for LLMs.

02Architecture
Meta-Transformer

Meta-Transformer

A unified framework performing learning across 12 different modalities with a shared backbone.

03Architecture
Retrieve In-Context Examples for LLMs

Retrieve In-Context Examples for LLMs

A framework to iteratively train dense retrievers that identify high-quality in-context examples.

04Retrieval
FLASK

FLASK

Proposes fine-grained evaluation of LLMs decomposed into 12 alignment skill sets.

05Evaluation
CM3Leon

CM3Leon

Meta's retrieval-augmented multi-modal language model that generates both text and images.

06Multimodal
Claude 2

Claude 2

Anthropic's second-generation LLM with a detailed model card on safety, alignment, and capabilities.

07Safety
Secrets of RLHF in LLMs

Secrets of RLHF in LLMs

A deep investigation into RLHF with a focus on the inner workings of PPO, including open-source code.

08Reinforcement Learning
LongLLaMA

LongLLaMA

Extends LLaMA's context length using a contrastive training process that reshapes the (key, value) space.

09Memory
Patch n' Pack: NaViT

Patch n' Pack: NaViT

A vision transformer handling any aspect ratio and resolution through sequence packing.

10Architecture
LLMs as General Pattern Machines

LLMs as General Pattern Machines

Demonstrates LLMs serve as general sequence modelers without additional training.

11Reasoning
HyperDreamBooth

HyperDreamBooth

A smaller, faster, and more efficient version of DreamBooth for personalizing text-to-image models.

12Multimodal
Teaching Arithmetic to Small Transformers

Teaching Arithmetic to Small Transformers

Trains small transformers on chain-of-thought style data for arithmetic with large gains.

13Reasoning
AnimateDiff

AnimateDiff

Animates frozen text-to-image diffusion models via a plug-in motion modeling module.

14Multimodal
Generative Pretraining in Multimodality (Emu)

Generative Pretraining in Multimodality (Emu)

A transformer-based multimodal foundation model for generating images and text.

15Multimodal
A Survey on Evaluation of LLMs

A Survey on Evaluation of LLMs

A comprehensive overview of evaluation methods covering what, where, and how to evaluate LLMs.

16Evaluation
How Language Models Use Long Contexts (Lost-in-the-Middle)

How Language Models Use Long Contexts (Lost-in-the-Middle)

Shows LLM performance drops when relevant information is in the middle of a long context.

17Memory
LLMs as Effective Text Rankers

LLMs as Effective Text Rankers

A prompting technique that enables open-source LLMs to perform SOTA text ranking.

18Retrieval
Multimodal Generation with Frozen LLMs

Multimodal Generation with Frozen LLMs

Maps images to LLM token space enabling models like PaLM and GPT-4 to handle visual tasks without parameter updates.

19Multimodal
CodeGen2.5

CodeGen2.5

Salesforce's new 7B code LLM trained on 1.5T tokens and optimized for fast sampling.

20Code
Elastic Decision Transformer

Elastic Decision Transformer

An advance over Decision Transformers that enables trajectory stitching at inference time.

21Reinforcement Learning
Robots That Ask for Help

Robots That Ask for Help

A framework for calibrating LLM-based robot planners so they ask for help when uncertain.

22Robotics
Physics-based Motion Retargeting in Real-Time

Physics-based Motion Retargeting in Real-Time

Uses RL to retarget motions from sparse human sensor data to characters of various morphologies.

23Multimodal
Scaling Transformer to 1 Billion Tokens (LongNet)

Scaling Transformer to 1 Billion Tokens (LongNet)

Microsoft's Transformer variant scaling sequence length past 1B tokens.

24Architecture
InterCode

InterCode

A framework treating interactive coding as a reinforcement learning environment.

25Reinforcement Learning
182 weeks of AI research · papers per week
Week of Sep 28–Oct 4, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Apr 2023Hover a week to inspect · select to openSep 2026