Tag: Mobile ML
-
TFLite Model Conversion: 10 Commands That Actually Work
Ten TFLite conversion commands that actually work in production โ with the edge cases, quantization tradeoffs, and debugging tricks the docs skip.
-
TFLite & ONNX Mobile Setup: 2x Speed Left on Table
Default TFLite and ONNX configs waste 2-8x performance. Here's the exact setupโdelegates, threads, quantizationโthat closes the gap on ARM devices.
-
ONNX Runtime Mobile: 8ms Inference on iPhone 13
Cut mobile inference from 200ms to 8ms by switching to ONNX Runtime. Benchmarks, CoreML quirks, and when TFLite still wins.
-
Edge AI Object Detection: YOLO Mobile Optimization Guide
YOLOv5n at 320px beats over-engineered quantization for mobile MVPs. Preprocessing tricks and NMS optimization that actually ship.
-
Real-time Whisper Is a Battery Nightmare (Here’s How to Fix It)
Real-time Whisper drains 1% battery/min. VAD + adaptive inference + thermal throttling bring it down to 0.2%. Benchmarks on iPhone 13 Pro.