🚀NEW LABGetting Started with Claude AgentsStart lab
DAIR.AI · Curated weekly since April 2023

AI Papers of the Week

Every paper worth reading in AI, hand-picked one week at a time.

2,650
Papers
182
Weekly issues
2023
Since

Discover and explore top AI papers with Claude Code or Codex

npx @dair-ai/mcp setup
392 papers · TrainingClear filters →
LongLoRA

LongLoRA

An efficient LoRA-based fine-tuning recipe for extending LLM context windows without expensive full fine-tuning.

337Training
Struc-Bench (LLMs for Structured Data)

Struc-Bench (LLMs for Structured Data)

Studies how LLMs handle complex structured-data generation and proposes a structure-aware fine-tuning method.

338Training
Textbooks Are All You Need II (phi-1.5)

Textbooks Are All You Need II (phi-1.5)

Microsoft's phi-1.5 demonstrates that a 1.3B model trained on "textbook-quality" synthetic data rivals much larger models on reasoning.

339Data
Radiology-Llama 2

Radiology-Llama 2

A Llama 2-based LLM specialized for radiology report generation.

340Training
Transformers as Support Vector Machines

Transformers as Support Vector Machines

A theoretical paper establishing a formal connection between self-attention optimization and hard-margin SVM problems.

341Training
Explaining Grokking

Explaining Grokking

DeepMind advances our understanding of grokking, predicting and confirming two novel phenomena that test their theory.

342Safety
FLM-101B

FLM-101B

A 101B parameter open LLM trainable on a $100K budget through a growth-based training strategy.

343Training
SAM-Med2D

SAM-Med2D

Adapts the Segment Anything Model (SAM) to 2D medical imaging through large-scale medical fine-tuning.

344Training
Survey on Instruction Tuning for LLMs

Survey on Instruction Tuning for LLMs

A comprehensive survey of instruction tuning covering methodology, dataset construction, and applications.

345Training
Prompt2Model

Prompt2Model

CMU's Prompt2Model automates the path from a natural-language task description to a deployable small special-purpose model.

346Agents
Platypus

Platypus

Platypus is a family of fine-tuned and merged LLMs that topped the Open LLM Leaderboard in August 2023.

347Training
Teach LLMs to Personalize

Teach LLMs to Personalize

A multitask-learning approach for personalized text generation without relying on predefined user attributes.

348Training
Synthetic Data Reduces Sycophancy

Synthetic Data Reduces Sycophancy

Google shows that fine-tuning on simple synthetic data can significantly reduce LLM sycophancy.

349Data
AutoRobotics-Zero

AutoRobotics-Zero

Discovers zero-shot adaptable robot policies from scratch, including the automatic discovery of Python control code.

350Robotics
Foundation Models in Vision

Foundation Models in Vision

A comprehensive survey on foundational models for computer vision and their open research directions.

351Multimodal
LoraHub

LoraHub

Enables efficient cross-task generalization via dynamic LoRA composition.

352Training
Llama 2

Llama 2

Meta's open-weight foundation model family with chat-tuned variants ranging from 7B to 70B parameters.

353Training
CM3Leon

CM3Leon

Meta's retrieval-augmented multi-modal language model that generates both text and images.

354Multimodal
Generative Pretraining in Multimodality (Emu)

Generative Pretraining in Multimodality (Emu)

A transformer-based multimodal foundation model for generating images and text.

355Multimodal
CodeGen2.5

CodeGen2.5

Salesforce's new 7B code LLM trained on 1.5T tokens and optimized for fast sampling.

356Code
Extending Context Window of LLMs (PI)

Extending Context Window of LLMs (PI)

Position Interpolation extends LLaMA's context to 32K with minimal fine-tuning (within 1000 steps).

357Memory
Visual Navigation Transformer (ViNT)

Visual Navigation Transformer (ViNT)

A foundation model for vision-based robotic navigation built on flexible Transformers.

358Robotics
Textbooks Are All You Need (phi-1)

Textbooks Are All You Need (phi-1)

Introduces a 1.3B parameter code LLM trained on textbook-quality data.

359Data
ClinicalGPT

ClinicalGPT

A language model optimized through extensive and diverse medical data and multi-turn dialogue.

360Training
182 weeks of AI research · papers per week
Week of Sep 28–Oct 4, 202610 papers →
Apr2023
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2024
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2025
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Jan2026
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Apr 2023Hover a week to inspect · select to openSep 2026