Hi,
Below is a similar topic for your reference:
Could you test with --dumpLayerInfo --dumpProfile --profilingVerbosity=detailed --separateProfileRun --useCudaGraph --noDataTransfers argument and share the output with us?
Thanks.
Hi,
Below is a similar topic for your reference:
Could you test with --dumpLayerInfo --dumpProfile --profilingVerbosity=detailed --separateProfileRun --useCudaGraph --noDataTransfers argument and share the output with us?
Thanks.
| Topic | Replies | Views | Activity | |
|---|---|---|---|---|
| Jetson Thor - INT8 quantization show no performance gain over FP16 | 7 | 533 | January 26, 2026 | |
| Jetson Thor - INT8 quantization show no performance gain over FP16 (2) | 5 | 463 | March 27, 2026 | |
| Int8 TensorCores for Jetson | 6 | 1522 | April 26, 2023 | |
| INT8 throughput and latency worse than FP16 for MiDas DPT Hybrid model on Thor | 3 | 240 | January 6, 2026 | |
| Inference Speed | 5 | 1120 | March 30, 2023 | |
| Int8 is not faster than fp16 on xavier | 4 | 901 | August 28, 2020 | |
| The NVIDIA DRIVE AGX Thor does not benefit from INT8 quantization, while other platforms do | 10 | 266 | February 26, 2026 | |
| Low ViT Performance Gain on Jetson Thor Using FP8 vs FP16 | 12 | 867 | February 13, 2026 | |
| How to verify if QAT TRT engine is indeed INT8 on Xavier | 15 | 820 | September 22, 2022 | |
| Pytorch network -> onnx -> tensorrt performance(run frequency) question | 6 | 621 | December 7, 2023 |