Efficient AI and edge inference: Artificial Intelligence course | Zoonk
58. Efficient AI and edge inference
Make models smaller, faster, and cheaper with quantization, distillation, pruning, caching, and edge deployment. You will decide when to run AI on a server, browser, phone, or embedded device.