Example: tourism industry
8-bit Inference with TensorRT - NVIDIA

8-bit Inference with TensorRT - NVIDIA

Back to document page

May 8, 2017. Intro Goal: Convert FP32 CNNs into INT8 without significant accuracy loss. Why: INT8 math has higher throughput, and lower memory requirements. Challenge: INT8 has significantly lower precision and dynamic range than FP32.

  2017, Tensorrt

Download 8-bit Inference with TensorRT - NVIDIA


Information

Domain:

Source:

Link to this page:

Please notify us if you found a problem with this document:

Other abuse

Advertisement

Related search queries