Knowledge Distillation — Roadmap Step & Resources
Training a smaller 'student' model to mimic a larger 'teacher' model
Open on Cached Info