arXiv · 1812.00249
On Compressing U-net Using Knowledge Distillation
Abstract
We study the use of knowledge distillation to compress the U-net architecture. We show that, while standard distillation is not sufficient to reliably train a compressed U-net, introducing other regularization methods, such as batch normalization and class re-weighting, in knowledge distillation significantly improves the training process. This allows us to compress a U-net by over 1000x, i.e., to 0.1% of its original number of parameters, at a negligible decrease in performance.
Explore related subjects
Keep this discovery
Karttikeya Mangalam, Mathieu Salzamann. 2018-12-01. On Compressing U-net Using Knowledge Distillation. https://arxiv.org/abs/1812.00249
Cite the original work for its findings. Save a collection to share your selection of sources.