🚀NEW LABGetting Started with Claude AgentsStart lab
Data

Mitigating Memorization in LLMs

First page
Mitigating Memorization in LLMs
Paper summary

presents a modification of the next-token prediction objective called goldfish loss to help mitigate the verbatim generation of memorized training data; it uses a simple technique that excludes a pseudorandom subset of training tokens at training time; they show that the goldfish loss resists memorization and keeps the model useful; however, it may need to train for longer to more effectively learn from the training data.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack