🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023Issue 180 · Sep 14 – Sep 20, 2026

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

2,315
Papers
180
Weekly issues
2023
Since
This week · 10 papersView the full issue →
Challenges & Application of LLMs

Challenges & Application of LLMs

A comprehensive enumeration of open challenges and application domains for LLMs.

02Safety
Retentive Network (RetNet)

Retentive Network (RetNet)

Microsoft's proposed foundation architecture aiming to replace Transformer attention for LLMs.

03Architecture
Meta-Transformer

Meta-Transformer

A unified framework performing learning across 12 different modalities with a shared backbone.

04Architecture
Retrieve In-Context Examples for LLMs

Retrieve In-Context Examples for LLMs

A framework to iteratively train dense retrievers that identify high-quality in-context examples.

05Retrieval
FLASK

FLASK

Proposes fine-grained evaluation of LLMs decomposed into 12 alignment skill sets.

06Evaluation
CM3Leon

CM3Leon

Meta's retrieval-augmented multi-modal language model that generates both text and images.

07Multimodal
Claude 2

Claude 2

Anthropic's second-generation LLM with a detailed model card on safety, alignment, and capabilities.

08Safety
Secrets of RLHF in LLMs

Secrets of RLHF in LLMs

A deep investigation into RLHF with a focus on the inner workings of PPO, including open-source code.

09Reinforcement Learning
LongLLaMA

LongLLaMA

Extends LLaMA's context length using a contrastive training process that reshapes the (key, value) space.

10Memory
Patch n' Pack: NaViT

Patch n' Pack: NaViT

A vision transformer handling any aspect ratio and resolution through sequence packing.

11Architecture
LLMs as General Pattern Machines

LLMs as General Pattern Machines

Demonstrates LLMs serve as general sequence modelers without additional training.

12Reasoning
HyperDreamBooth

HyperDreamBooth

A smaller, faster, and more efficient version of DreamBooth for personalizing text-to-image models.

13Multimodal
Teaching Arithmetic to Small Transformers

Teaching Arithmetic to Small Transformers

Trains small transformers on chain-of-thought style data for arithmetic with large gains.

14Reasoning
AnimateDiff

AnimateDiff

Animates frozen text-to-image diffusion models via a plug-in motion modeling module.

15Multimodal
Generative Pretraining in Multimodality (Emu)

Generative Pretraining in Multimodality (Emu)

A transformer-based multimodal foundation model for generating images and text.

16Multimodal
A Survey on Evaluation of LLMs

A Survey on Evaluation of LLMs

A comprehensive overview of evaluation methods covering what, where, and how to evaluate LLMs.

17Evaluation
How Language Models Use Long Contexts (Lost-in-the-Middle)

How Language Models Use Long Contexts (Lost-in-the-Middle)

Shows LLM performance drops when relevant information is in the middle of a long context.

18Memory
LLMs as Effective Text Rankers

LLMs as Effective Text Rankers

A prompting technique that enables open-source LLMs to perform SOTA text ranking.

19Retrieval
Multimodal Generation with Frozen LLMs

Multimodal Generation with Frozen LLMs

Maps images to LLM token space enabling models like PaLM and GPT-4 to handle visual tasks without parameter updates.

20Multimodal
CodeGen2.5

CodeGen2.5

Salesforce's new 7B code LLM trained on 1.5T tokens and optimized for fast sampling.

21Code
Elastic Decision Transformer

Elastic Decision Transformer

An advance over Decision Transformers that enables trajectory stitching at inference time.

22Reinforcement Learning
Robots That Ask for Help

Robots That Ask for Help

A framework for calibrating LLM-based robot planners so they ask for help when uncertain.

23Robotics
Physics-based Motion Retargeting in Real-Time

Physics-based Motion Retargeting in Real-Time

Uses RL to retarget motions from sparse human sensor data to characters of various morphologies.

24Multimodal
Scaling Transformer to 1 Billion Tokens (LongNet)

Scaling Transformer to 1 Billion Tokens (LongNet)

Microsoft's Transformer variant scaling sequence length past 1B tokens.

25Architecture
180 weeks of AI research · papers per week
Week of Sep 14–20, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Apr 2023Hover a week to inspect · select to openSep 2026