No AI summary available for this article.
Why It Matters
Knowledge distillation offers an efficient route to transfer a task-adapted vision-language teacher to a compact student.
Provenance
Discovered via ArXiv and published by ArXiv.
Key Claims
Original description
Knowledge distillation offers an efficient route to transfer a task-adapted vision-language teacher to a compact student. The training target in current vision-language distillation methods is typically constructed from the teacher prediction and applied uniformly to all training samples, making it unreliable under class and domain shifts. In this paper, we argue that distillation target construction should be treated as a dynamic training decision rather than a fixed recipe. To this end, we propose OnPoKD, an on-policy distillation framework for vision-language model adaptation. To the best o...
Discovered via ArXiv
Research papers and preprints from arXiv.
Publisher: arxiv.org
ID: http://arxiv.org/abs/2609.10321v1 · Indexed about 2 hours ago