Tag: TensorRT
-
Raspberry Pi 5 vs Jetson Nano: Budget Edge AI Latency Test
Pi 5 vs Jetson Nano latency test: INT8 models close the 38ms gap to 4.4ms. Real power draw, conversion pain, and when CUDA actually matters.
-
RT-2 vs OpenVLA: Inference Speed on Jetson AGX Orin
RT-2 runs in 312ms on Jetson Orin with TensorRT. OpenVLA hits 1.2s and won't export to ONNX yet. Real latency numbers for edge VLA deployment.
-
YOLOv8 INT8 Quantization: 4x Faster on Jetson Orin
INT8 quantization pushed YOLOv8 from 45 to 180 FPS on Jetson Orin Nano. PTQ beat QAT with 10 lines of calibration code. Here's the recipe.
-
Edge AI vs Cloud AI: Architecture Guide for Factories
Edge AI vs Cloud AI for factories: latency, cost, and reliability compared across 3 real deployments. The hybrid approach won.
-
ONNX Runtime๊ณผ TensorRT๋ก ์ปดํจํฐ๋น์ ๋ชจ๋ธ ๊ฒฝ๋ํํ๊ธฐ: Edge AI ๋ฐฐํฌ ์ค์ ๊ฐ์ด๋
Speed up YOLO inference 8x on Jetson with TensorRT INT8 quantization. ONNX conversion pitfalls and FP16 vs INT8 accuracy tradeoffs included.