
Hallucination in LLMs Survey
A comprehensive survey of hallucination in LLMs, covering taxonomy, causes, evaluation, and mitigation.

Simplifying Transformer Blocks
Researchers show that many components of the standard transformer block can be removed with no loss in training speed or quality.

In-Context Learning Generalization Limits
Investigates whether transformers' in-context learning can generalize beyond the distribution of their pretraining data.

MusicGen
Meta's MusicGen is a single-stage transformer LLM for music generation that operates over compressed discrete audio tokens.

AltUp (Alternating Updates)
Google's AltUp lets transformers benefit from wider representations without paying the full compute cost at every layer.

Rephrase and Respond (RaR)
An effective prompting method where the LLM rephrases and expands the user's question before answering it.

On the Road with GPT-4V
An exhaustive evaluation of GPT-4V applied to autonomous driving scenarios.

GPT4All Technical Report
The GPT4All technical report documents the model family and the open ecosystem built around democratizing local LLMs.

S-LoRA
S-LoRA enables serving thousands of LoRA adapters concurrently on a single GPU through memory-paging and custom CUDA kernels.

FreshLLMs (FreshQA)
Introduces FreshQA, a dynamic benchmark designed to stress-test LLMs on time-sensitive knowledge.
Get next week’s papers.
The same picks and the same summaries, in your inbox. Free, and 176 issues deep.
Subscribe on Substack