7 ms·On-device CPU inference is the real flex here. Optimization probably mattered as much as modeling.by 1ilit 7mo agoOn-device CPU inference is the real flex here. Optimization probably mattered as much as modeling.