🚀NEW LABGetting Started with Claude AgentsStart lab
Safety

Contextual Hallucinations Mitigation in LLMs

First page
Contextual Hallucinations Mitigation in LLMs
Paper summary

proposes a new method that detects and significantly reduces contextual hallucinations in LLMs (e.g., reduces by 10% in the XSum summarization task); builds a hallucination detection model based on input features given by the ratio of attention weights on the context vs. newly generated tokens (for each attention head); the hypothesis is that contextual hallucinations are related to the extent to which an LLM attends to the provided contextual information; they also propose a decoding strategy based on their detection method which mitigates the contextual hallucination; the detector can also be transferred across models without the need for retraining.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack