Tag: ONNX Runtime
-
ONNX Runtime๊ณผ TensorRT๋ก ์ปดํจํฐ๋น์ ๋ชจ๋ธ ๊ฒฝ๋ํํ๊ธฐ: Edge AI ๋ฐฐํฌ ์ค์ ๊ฐ์ด๋
Speed up YOLO inference 8x on Jetson with TensorRT INT8 quantization. ONNX conversion pitfalls and FP16 vs INT8 accuracy tradeoffs included.