Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs

Dec-24-2025, 15:00:42 GMT–Neural Information Processing Systems

To mitigate memorization, we introduce a subtle modification to the next-token training objective that we call the goldfish loss. During training, a randomly sampled subsets of tokens are excluded from the loss computation. These dropped tokens are not memorized by the model, which prevents verbatim reproduction of a complete chain of tokens from the training set. We run extensive experiments training billion-scale LLaMA-2 models, both pre-trained and trained from scratch, and demonstrate significant reductions in extractable memorization with little to no impact on downstream benchmarks.

large language model, machine learning, natural language, (7 more...)

Neural Information Processing Systems

Dec-24-2025, 15:00:42 GMT

Conferences Web Page

Add feedback

Technology:
- Information Technology > Artificial Intelligence
  - Natural Language > Large Language Model (0.96)
  - Machine Learning > Memory-Based Learning
    - Rote Learning (0.92)