Antidistillation Sampling

Savani, Yash, Trockman, Asher, Feng, Zhili, Xu, Yixuan Even, Schwarzschild, Avi, Robey, Alexander, Finzi, Marc, Kolter, J. Zico

arXiv.org Artificial Intelligence 

Frontier models that generate extended reasoning traces inadvertently produce rich token sequences that can facilitate model distillation. Recognizing this vulnerability, model owners may seek sampling strategies that limit the effectiveness of distillation without compromising model performance. Antidistillation sampling provides exactly this capability. By strategically modifying a model's next-token probability distribution, antidistillation sampling poisons reasoning traces, rendering them significantly less effective for distillation while preserving the model's practical utility. For further details, see https://antidistillation.com.

Duplicate Docs Excel Report

Title
None found

Similar Docs  Excel Report  more

TitleSimilaritySource
None found