
Universal Adversarial LLM Attacks
Finds universal and transferable adversarial attacks that cause aligned models like ChatGPT and Bard to generate objectionable behaviors.

RT-2
Google DeepMind's end-to-end vision-language-action model that learns from both web and robotics data to control robots.

Med-PaLM Multimodal
Introduces a generalist biomedical AI system and a new multimodal biomedical benchmark with 14 tasks.

Tracking Anything in High Quality
A framework for high-quality tracking-anything in videos combining segmentation and refinement.

Foundation Models in Vision
A comprehensive survey on foundational models for computer vision and their open research directions.

L-Eval
A standardized evaluation suite for long-context language models.

LoraHub
Enables efficient cross-task generalization via dynamic LoRA composition.

Survey of Aligned LLMs
A comprehensive overview of alignment approaches covering data, training, and evaluation.

WavJourney
Leverages LLMs to orchestrate audio generation models for compositional storytelling.

FacTool
A task- and domain-agnostic framework for factuality detection of LLM-generated text.
Get next week’s papers.
The same picks and the same summaries, in your inbox. Free, and 176 issues deep.
Subscribe on Substack