Quick Specs Small form-factor, 75-Watt design fits any scale-out server. INT8 operations slash latency by 15X. Hardware-decode engine capable of transcoding and inferencing 35 HD video streams in real time. GPUArchitecture NVIDIA Pascal™ Single-Precision Performance 5.5 TeraFLOPS* Integer Operations (INT8) 22 TOPS* (TeraOperations per Second) GPU Memory 8 GB Memory Bandwidth 192 GB/s System Interface Low-Profile PCI Express Form Factor Max Power 75W Enhanced Programmability with Page Migration Engine Yes ECC Protection Yes Server-Optimized for Data Center Deployment Yes Hardware-Accelerated Video Engine 1x Decode Engine, 2x Encode Engine The NVIDIA Tesla P4 is powered by the revolutionary NVIDIA Pascal™ architecture and purpose-built to boost efficiency for scale-out servers running deep learning workloads, enabling smart responsive AI-based services. It slashes inference latency by 15X in any hyperscale infrastructure and provides an incredible 60X better energy efficiency than CPUs. This unlocks a new wave of AI services previous impossible due to latency limitations.