NVIDIA L4 Tensor Core GPU
NVIDIA
GPUs
AI and ML
The NVIDIA L4 Tensor Core GPU delivers powerful acceleration for AI inference, video analytics, and graphics rendering, all in a low-profile, low-power PCIe form factor. It’s the ideal solution for modern workloads across cloud, enterprise, and edge data centers. Up to 4× faster inference throughput than the previous NVIDIA T4.
Compact yet capable: up to 120× better AI inference performance than CPU-only systems.
Why Choose NVIDIA L4?
- Hyperscalers deploying multi-instance inferenc
- Enterprises needing high-efficiency AI servers
- System builders delivering compact, low-power solutions
- CSPs offering GPU-accelerated VDI & graphics
- Energy Efficient 72W TDP makes it easy to deploy at scale
- Flexible Drop-in PCIe design, compatible with 1U servers
- Smart NVIDIA AI software stack (TensorRT, Triton, CUDA, etc.) fully supported
- Video-Optimized – Best-in-class AV1 encoding for media workloads
Key Differentiators:
- Capability NVIDIA T4 NVIDIA L4
- GPU architecture Turing Ada Lovelace
- Tensor Cores 320 (older gen) 232 (4th Gen)
- FP8 Support ✘ ✔
- Video Encode H.264/HEVC H.264/HEVC + AV!
- Memory 16 GB 24 GB
- Power 70W 72W
Data sheet(s)