Quiz 31 Question 12 of 20

After completing deep learning model training, you want to compile and optimize the neural network specifically to run with maximum throughput and minimum latency on NVIDIA GPU-based inference servers. Which NVIDIA SDK should you use to convert the model into an optimized runtime engine?

Select an answer to reveal the explanation.

Motivation