Knowledge Distillation: Compressing Large Models

Knowledge Distillation: Compressing Large Models

Knowledge distillation trains smaller student models to mimic larger teacher models.

Related Chronicles: The Compression Catastrophe (2035)