acceptodds
Under review as a conference paper at ICLR 2027

Distillation Lineage Inspector: Black-Box Auditing of Model Distillation in LLMs

Abstract

Model distillation has emerged as a widely used technique for creating efficient models tailored to specific tasks or domains. However, its reliance on knowledge from foundation models raises significant legal concerns regarding intellectual property rights. To address this issue, we propose the Distillation Lineage Inspector (DLI) framework, which enables model developers to determine whether their large language models (LLMs) have been distilled without authorization, even in black-box settings where training data and model architecture are inaccessible. DLI is effective across both open-source and closed-source LLMs. Experiments show that DLI achieves 80% accuracy with as few as 10 prompts in fully black-box settings and yields a 45% improvement in accuracy over the best baseline under standard experimental conditions. Furthermore, we analyze how auditor knowledge of target models influences performance, providing practical insight for building privacy-preserving and regulation-compliant AI systems.

Then back it, or bet against it.

Related papers

Open the market on this paper to see 7 more related papers.