Large Language Model Unlearning via Embedding-Corrupted Prompts

Oct-10-2025, 17:55:43 GMT–Neural Information Processing Systems

Instead of relying on the LLM itself to unlearn, we enforce an unlearned state during inference by employing a prompt classifier to identify and safeguard prompts to forget.

arxiv preprint arxiv, classifier, dataset, (14 more...)

Neural Information Processing Systems

Oct-10-2025, 17:55:43 GMT

Conferences PDF

Add feedback

Country:
- South America > Chile
  - Santiago Metropolitan Region > Santiago Province > Santiago (0.04)
- North America > United States
  - Virginia (0.04)
  - Massachusetts (0.04)
  - California
    - Santa Cruz County > Santa Cruz (0.04)
    - San Diego County > San Diego (0.04)
- Europe
  - Montenegro (0.04)
  - Poland (0.04)
  - Central Europe (0.04)
  - Italy > Calabria
    - Catanzaro Province > Catanzaro (0.04)
- Asia
  - Kazakhstan (0.04)
  - China (0.04)
  - Middle East > Kuwait
    - Capital Governorate > Kuwait City (0.04)
  - India > Maharashtra
    - Mumbai (0.04)

Genre:
- Research Report > Experimental Study (1.00)

Industry:
- Media (1.00)
- Leisure & Entertainment (1.00)
- Law (1.00)
- Information Technology > Security & Privacy (1.00)
- Government (1.00)
- Education (1.00)
- Health & Medicine > Therapeutic Area
  - Psychiatry/Psychology (0.45)

Technology:
- Information Technology > Artificial Intelligence
  - Natural Language
    - Large Language Model (1.00)
    - Chatbot (1.00)
  - Machine Learning
    - Neural Networks > Deep Learning (1.00)
    - Performance Analysis > Accuracy (0.95)

Duplicate Docs Excel Report

Title
Large Language Model Unlearning via Embedding-Corrupted Prompts

Similar Docs Excel Report more

Title	Similarity	Source
None found