Transcription of NVIDIA AMPERE GA102 GPU ARCHITECTURE
{{id}} {{{paragraph}}}
NVIDIA AMPERE GA102 GPU ARCHITECTURE Second-Generation RTX NVIDIA AMPERE GA102 GPU ARCHITECTURE ii Table of Contents Introduction 5 GA102 Key Features 7 2x FP32 Processing 7 Second-Generation RT Core 7 Third-Generation Tensor Cores 8 GDDR6X and GDDR6 Memory 8 Third-Generation NVLink 8 PCIe Gen 4 9 AMPERE GPU ARCHITECTURE In-Depth 10 GPC, TPC, and SM High-Level ARCHITECTURE 10 ROP Optimizations 11 GA10x SM ARCHITECTURE 11 2x FP32 Throughput 12 Larger and Faster Unified Shared Memory and L1 Data Cache 13 Performance Per Watt 16 Second-Generation Ray Tracing Engine in GA10x GPUs 17 AMPERE ARCHITECTURE RTX Processors in Action 19 GA10x GPU Hardware Acceleration for Ray-Traced Motion Blur 20 Third-Generation Tensor Cores in GA10x GPUs 24 Comparison of Turing vs GA10x GPU Tensor Cores 24 NVIDIA AMPERE ARCHITECTURE Tensor Cores Support New DL Data Types 26 Fine-Grained Structured Sparsity 26 NVIDIA DLSS 8K 28 GDDR6X Memory 30 RTX IO 32 Introducing NVIDIA RTX IO 33 How NVIDIA RTX IO Works 34 Display and Video Engine 38 DisplayPort with DSC 38 HDMI with DSC 38 Fifth Generation NVDEC - Hardware-Accelerated Video Decoding 39 AV1 Hardware Decode 40 Seventh Generation
Programmable Shading Cores, which consist of NVIDIA CUDA Cores RT Cores, which accelerate Bounding Volume Hierarchy (BVH) traversal and intersection of scene geometry during ray tracing Tensor Cores, which provide enormous …
Domain:
Source:
Link to this page:
Please notify us if you found a problem with this document:
{{id}} {{{paragraph}}}