🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

1,760
Papers
176
Weekly issues
2023
Since
618 papers · TrainingClear filters →
Unifying LLMs & Knowledge Graphs

Unifying LLMs & Knowledge Graphs

A roadmap for combining LLMs with knowledge graphs for stronger reasoning.

577Retrieval
Fine-Grained RLHF

Fine-Grained RLHF

Trains LMs with segment-level human feedback rather than whole-response preferences.

578Reinforcement Learning
Hierarchical Vision Transformer (Hiera)

Hierarchical Vision Transformer (Hiera)

Pretrains ViTs with MAE while removing unnecessary multi-stage complexity.

579Architecture
Imitating Reasoning Process of Larger LLMs (Orca)

Imitating Reasoning Process of Larger LLMs (Orca)

Microsoft's 13B model that imitates GPT-4's reasoning traces.

580Reasoning
Fine-Tuning Language Models with Just Forward Passes (MeZO)

Fine-Tuning Language Models with Just Forward Passes (MeZO)

A memory-efficient zeroth-order optimizer for LLM fine-tuning.

581Training
MERT

MERT

An acoustic music understanding model with large-scale self-supervised training.

582Multimodal
Bytes Are All You Need

Bytes Are All You Need

Performs classification directly on file bytes without decoding.

583Training
Direct Preference Optimization (DPO)

Direct Preference Optimization (DPO)

Rafailov et al.'s simpler alternative to RLHF that rivals full RL-based alignment.

584Reinforcement Learning
SQL-PaLM

SQL-PaLM

An LLM-based Text-to-SQL system built on PaLM-2.

585Code
CodeTF

CodeTF

An open-source Transformer library for state-of-the-art code LLMs.

586Code
QLoRA

QLoRA

Tim Dettmers' breakthrough technique enabling 65B LLM fine-tuning on a single 48GB GPU.

587Training
LIMA

LIMA

Meta's 65B LLaMA fine-tuned on just 1,000 curated examples - showing alignment needs less data than believed.

588Training
Gorilla

Gorilla

A fine-tuned LLaMA-based model that surpasses GPT-4 on API call generation.

589Agents
The False Promise of Imitating Proprietary LLMs

The False Promise of Imitating Proprietary LLMs

Berkeley's critical analysis of open-source imitation of proprietary LLMs.

590Training
Sophia

Sophia

A simple, scalable second-order optimizer with negligible per-step overhead.

591Efficiency
Reinventing RNNs for the Transformer Era (RWKV)

Reinventing RNNs for the Transformer Era (RWKV)

Combines parallelizable training of Transformers with efficient RNN inference.

592Architecture
DoReMi

DoReMi

Optimizes data mixtures for faster language model pretraining.

593Training
CodeT5+

CodeT5+

An open code LLM family for code understanding and generation.

594Code
Symbol tuning

Symbol tuning

Fine-tunes LMs on in-context input-label pairs with natural-language labels replaced by arbitrary symbols.

595Reasoning
Incidental Bilingualism in PaLM's Translation Capability

Incidental Bilingualism in PaLM's Translation Capability

Explores where PaLM's translation ability actually comes from.

596Training
InstructBLIP

InstructBLIP

Visual-language instruction tuning built on BLIP-2.

597Multimodal
MultiModal-GPT

MultiModal-GPT

A vision-language model for multi-round dialogue fine-tuned from OpenFlamingo.

598Multimodal
scGPT

scGPT

A foundation model for single-cell multi-omics pretrained on 10 million cells.

599Training
GPTutor

GPTutor

A ChatGPT-powered VSCode extension for code explanation.

600Training
176 weeks of AI research · papers per week
Week of Aug 17–23, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Apr 2023Hover a week to inspect · select to openAug 2026