
Tracking Everything Everywhere All at Once (OmniMotion)
Test-time optimization for dense, long-range motion estimation.

AlphaDev
DeepMind's deep RL agent discovering faster sorting algorithms from scratch, now in LLVM.

Sparse-Quantized Representation (SpQR)
Tim Dettmers' near-lossless LLM compression technique.

MusicGen
A simple and controllable model for music generation using a single-stage Transformer.

Augmenting LLMs with Databases (ChatDB)
Combines an LLM with SQL databases as a symbolic memory framework.

Concept Scrubbing in LLM (LEACE)
Least-squares Concept Erasure - erases a target concept from every layer of a neural network.

Fine-Grained RLHF
Trains LMs with segment-level human feedback rather than whole-response preferences.

Hierarchical Vision Transformer (Hiera)
Pretrains ViTs with MAE while removing unnecessary multi-stage complexity.

Humor in ChatGPT
Explores ChatGPT's capabilities to grasp and reproduce humor.

Imitating Reasoning Process of Larger LLMs (Orca)
Microsoft's 13B model that imitates GPT-4's reasoning traces.
Get next week’s papers.
The same picks and the same summaries, in your inbox. Free, and 176 issues deep.
Subscribe on Substack