Skip to content
InfoResearchPreprintLLM-specific

Poster: A Preliminary Study of LLM Distillation Inference

Published
Record updated
View JSON

Summary

This poster presents a preliminary study of distillation inference, a method for determining whether a suspect model was distilled from a proprietary teacher LLM or trained independently. The approach frames the question as a hypothesis test, using shadow models trained on either teacher reasoning traces or reference answers to calibrate a p-value from how closely the suspect predicts the teacher's reasoning outputs. Using Qwen2.5-7B as the teacher and Llama-3.2-3B for the suspects, the test achieves a true positive rate of 1.0 at a significance level of 0.02.