Counterfactual Memorization in Neural Language Models
Zhang, Chiyuan, Ippolito, Daphne, Lee, Katherine, Jagielski, Matthew, Tramèr, Florian, Carlini, Nicholas
–arXiv.org Artificial Intelligence
Modern neural language models widely used in tasks across NLP risk memorizing sensitive information from their training data. As models continue to scale up in parameters, training data, and compute, understanding memorization in language models is both important from a learning-theoretical point of view, and is practically crucial in real world applications. An open question in previous studies of memorization in language models is how to filter out "common" memorization. In fact, most memorization criteria strongly correlate with the number of occurrences in the training set, capturing "common" memorization such as familiar phrases, public knowledge or templated texts. In this paper, we provide a principled perspective inspired by a taxonomy of human memory in Psychology. From this perspective, we formulate a notion of counterfactual memorization, which characterizes how a model's predictions change if a particular document is omitted during training. We identify and study counterfactually-memorized training examples in standard text datasets. We further estimate the influence of each training example on the validation set and on generated texts, and show that this can provide direct evidence of the source of memorization at test time.
arXiv.org Artificial Intelligence
Dec-23-2021
- Country:
- Asia
- Azerbaijan (0.68)
- India (1.00)
- Middle East > Israel (0.68)
- Europe
- North America > United States
- Maryland > Prince George's County (0.14)
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Virginia > Alleghany County (0.14)
- Washington > Snohomish County (0.14)
- Asia
- Genre:
- Research Report (1.00)
- Industry:
- Energy > Oil & Gas (1.00)
- Banking & Finance
- Education (1.00)
- Government
- Military > Navy (0.67)
- Regional Government
- Asia Government (0.92)
- Europe Government (0.67)
- North America Government > United States Government (1.00)
- Voting & Elections (0.92)
- Transportation > Air (0.93)
- Health & Medicine
- Consumer Health (0.93)
- Therapeutic Area (1.00)
- Leisure & Entertainment
- Information Technology > Security & Privacy (0.65)
- Consumer Products & Services (0.92)
- Law > Criminal Law (0.67)
- Materials > Chemicals
- Commodity Chemicals (0.67)
- Food & Agriculture > Agriculture (0.93)
- Law Enforcement & Public Safety > Crime Prevention & Enforcement (1.00)
- Technology: