Smart Devices, Sensing Hardware, and Embedded Interaction / Embedded Deployment, Automation Integration, and Device Constraints

Which compression techniques (e.g., quantization, pruning, distillation) most significantly affect on-device ML performance?

Similar questions

Related papers