NVIDIA L40

NVIDIA Components AI and ML, HPC, AI and ML

Unlock blazing visualization and AI power with the NVIDIA L40—Ada Lovelace-driven, petaflop-capable, and secure for 24/7 enterprise performance. It’s more than a GPU—it’s a multi-workload turbo-engine for the modern data centre.

L40 succeeds A40, The A40 (Ampere architecture) was designed as a data centre GPU for professional visualization, rendering, AI, and virtualization workloads. The L40 (Ada Lovelace architecture) takes over that role, bringing major improvements in ray tracing, Tensor Core performance (including FP8), memory bandwidth, and media engines (with AV1 support).

Why NVIDIA L40 Stands Out

  • Lightning-fast visuals & AI—Double the ray-tracing, FP8 inferencing, NVIDIA's newest architecture.
  • Built for enterprise—Rock-solid reliability, passive cooling, and data centre compliance.
  • Multi-role mastery—From virtual workstations and gaming-quality graphics to deep AI training and live video stream processing.
  • Future-ready flexibility—Perfect for cloud deployment, creative studios, engineering teams, and AI developers.
  • Exceptional Multi-Workload Capabilities
  • Versatile Performance: Built on NVIDIA’s Ada Lovelace architecture, the L40 delivers premium graphics, compute, AI, and virtualization performance across data centre workloads.
  • Double the Ray Tracing Power: With third-generation RT Cores, it achieves up to 2× the real-time ray-tracing performance of its predecessor—perfect for interactive rendering and photorealistic design workflows.

Supercharged Compute & AI Performance

  • Cutting-Edge Tensor Cores: The fourth-generation Tensor Cores with FP8 support deliver over 1 petaflop of inference performance, supercharging deep learning and AI pipelines.
  • Memory-Driven Throughput: Equipped with 48 GB GDDR6 ECC memory and wide 864 GB/s bandwidth, it handles complex data-intensive tasks with ease.
  • Engineered for Enterprise-Grade Reliability
  • Data Centre Ready: Designed for 24/7 operations, the L40 includes features like Secure Boot with internal root of trust, passive cooling, and NEBS Level 3 compliance—ensuring both security and endurance in critical environments.
  • Compact & Efficient Design: Packaged in a dual-slot, full-height (4.4" H × 10.5" L) form factor, it fits seamlessly into industry-standard servers and certified systems.

Scalability & Application Versatility

  • Ideal for Omniverse: As the backbone of NVIDIA Omniverse Enterprise, the L40 excels in powering XR, virtual reality, digital twins, and collaborative 3D environments—with accelerated ray-traced simulations and realistic synthetic rendering.
  • Seamless Virtualization: When paired with RTX Virtual Workstation (vWS) or vGPU software like vPC and vApps, it enables secure, high-performance virtual workstations across devices—empowering remote creative professionals.
  • Comprehensive AI & Data Workflows: Ideal for AI training, inference, data science, streaming, and content pipelines, with FP8-ready Tensor Cores and robust NVENC/NVDEC capabilities, including AV1 encoding/decoding.

Data sheet(s)

Vgpu L40 Datasheet