Model Distillation — Roadmap Step & Resources
Training a smaller model to mimic a larger one
Open on CachedInfo