Antidistillation Sampling

Yash Savani; Asher Trockman; Zhili Feng; Yixuan Even Xu; Avi Schwarzschild; Alexander Robey; Marc Finzi; J. Zico Kolter

Antidistillation Sampling

Artificial Intelligence 2025-10-28 v5 Computation and Language

Authors: Yash Savani , Asher Trockman , Zhili Feng , Yixuan Even Xu , Avi Schwarzschild , Alexander Robey , Marc Finzi , J. Zico Kolter

View on arXiv ↗ PDF ↗

Abstract

Frontier models that generate extended reasoning traces inadvertently produce rich token sequences that can facilitate model distillation. Recognizing this vulnerability, model owners may seek sampling strategies that limit the effectiveness of distillation without compromising model performance. Antidistillation sampling provides exactly this capability. By strategically modifying a model's next-token probability distribution, antidistillation sampling poisons reasoning traces, rendering them significantly less effective for distillation while preserving the model's practical utility. For further details, see https://antidistillation.com.

Keywords

knowledge distillation randomized algorithm machine learning

Cite

@article{arxiv.2504.13146,
  title  = {Antidistillation Sampling},
  author = {Yash Savani and Asher Trockman and Zhili Feng and Yixuan Even Xu and Avi Schwarzschild and Alexander Robey and Marc Finzi and J. Zico Kolter},
  journal= {arXiv preprint arXiv:2504.13146},
  year   = {2025}
}

Antidistillation Sampling

Abstract

Keywords

Cite

Related papers