
The Reversal Curse
Finds that LLMs trained on "A is B" fail to generalize to "B is A" - a surprisingly deep failure of learning.

Effective Long-Context Scaling (Meta)
Meta proposes a 70B long-context LLM that surpasses GPT-3.5-turbo-16k on long-context benchmarks.

Graph Neural Prompting (GNP)
A plug-and-play method that injects knowledge-graph information into frozen pretrained LLMs.

Vision Transformers Need Registers
Meta researchers identify artifact tokens in ViT feature maps and propose a trivial fix: add dedicated register tokens.

Boolformer
The first Transformer trained to perform end-to-end symbolic regression of Boolean functions.

LLaVA-RLHF
Adapts factually augmented RLHF to aligning large multimodal models, reducing hallucination without falling into reward-hacking pitfalls.

LLM Alignment Survey
A comprehensive survey of LLM alignment research spanning theoretical foundations to adversarial pressure.

Qwen
Alibaba releases the Qwen family of open LLMs with strong tool-use and planning capabilities for language agents.

MentaLLaMA
An open-source LLM family specialized for interpretable mental-health analysis on social media.

Logical Chain-of-Thought (LogiCoT)
A neurosymbolic framework that verifies and revises zero-shot CoT reasoning using symbolic-logic principles.
Get next week’s papers.
The same picks and the same summaries, in your inbox. Free, and 176 issues deep.
Subscribe on Substack