Description
If one has a trained neural network in Python Tensorflow and wants to run the C++ TensorRT engine in Windows, what are the options?
- Using TF-TRT on a Linux machine to generate the TensorRT model and then move the model to Windows to build the engine?
- Using the UFF model converter on a Linux machine to generate the TensorRT model and then move the model to Windows to build the engine?
- Convert to ONNX using a tool like tf2onnx and importing this ONNX into TensorRT on Windows?
- Build the network layer by layer in TensorRT?
Would any of the above work? Is there another method? What is the recommended method?
Environment
TensorRT Version: 7
GPU Type: P5000
CUDA Version: 10.0
CUDNN Version: 7.6.5
Operating System + Version: Windows 10
Python Version (if applicable): 3.7
TensorFlow Version (if applicable): 2.1