arXiv · 2109.06458
A Note on Knowledge Distillation Loss Function for Object Classification
Abstract
This research note provides a quick introduction to the knowledge distillation loss function used in object classification. In particular, we discuss its connection to a previously proposed logits matching loss function. We further treat knowledge distillation as a specific form of output regularization and demonstrate its connection to label smoothing and entropy-based regularization.
Explore related subjects
Keep this discovery
Defang Chen. 2021-09-14. A Note on Knowledge Distillation Loss Function for Object Classification. https://arxiv.org/abs/2109.06458
Cite the original work for its findings. Save a collection to share your selection of sources.