I am running ResNet50 onnx model using trtexec on Nvidia T4. They print Throughput and latency. I am confused why the throughput in QPS instead of Images/Sec. And also the numbers are extremely bad compared to MLPerf.
Related topics
| Topic | Replies | Views | Activity | |
|---|---|---|---|---|
| Inference Speed | 5 | 1140 | March 30, 2023 | |
| Trtexec performance not close to benchmarks | 1 | 583 | December 19, 2023 | |
| Orin nano/nx ResNet-50 benchmark on R36.4.3(jetpack6.2) | 7 | 732 | March 10, 2025 | |
| Same inference speed with Resnet50 for int8 and fp16 | 3 | 828 | November 30, 2020 | |
| An FCN with 360K parameters much slower compared to ResNet50 with 23M parameters | 3 | 481 | April 21, 2020 | |
| TensorRT has less batching throughput improvement than PyTorch on Jetson Nano | 2 | 535 | September 22, 2020 | |
| trtexec set input shape not working with | 2 | 5812 | August 5, 2021 | |
| Onnx -> TensorRT. No speed difference between models of different sizes | 5 | 974 | July 12, 2021 | |
| Unexpected TensorRT5.1.2 Results vs TRTIS1.0.0 Results | 0 | 787 | April 19, 2019 | |
| Why Orin's trtexec produces much better inference from onnx model than from engine model? | 6 | 731 | August 31, 2022 |