Hi,
There is a report that FP8 doesn’t have the expected performance with TensorRT on Thor:
We will check if the same perf regression issue for the INT8 case.
Thanks.
Hi,
There is a report that FP8 doesn’t have the expected performance with TensorRT on Thor:
We will check if the same perf regression issue for the INT8 case.
Thanks.
| Topic | Replies | Views | Activity | |
|---|---|---|---|---|
| Jetson Thor - INT8 quantization show no performance gain over FP16 (2) | 5 | 470 | March 27, 2026 | |
| Jetson Thor AGX - Poor INT8 performance | 6 | 526 | April 1, 2026 | |
| INT8 throughput and latency worse than FP16 for MiDas DPT Hybrid model on Thor | 3 | 244 | January 6, 2026 | |
| Post quantization aware training is slower than fp16 and post quantization | 12 | 2962 | September 25, 2024 | |
| [Hugging Face transformer models + pytorch_quantization] PTQ quantization int8 is slower than fp16 | 4 | 3221 | January 6, 2022 | |
| QAT int8 TRT engine slower than fp16 | 3 | 2570 | January 6, 2022 | |
| TensorRT --fp16 pre and post Int8 quantization | 1 | 196 | September 2, 2024 | |
| TRT Engin in INT8 is much slower than FP16 | 4 | 2159 | November 11, 2021 | |
| Same inference speed for INT8 and FP16 | 9 | 6489 | March 8, 2019 | |
| PeopleNet int8 shows small improvement over fp16 | 5 | 890 | July 1, 2021 |