Skip to product information
1 of 1

NVIDIA

NVIDIA 900-2G133-0080-000 | L40S 48GB Graphics Card, 350W, 4x DisplayPort 1.4a

SKU:900-2G133-0080-000

Stock Status: Enquire

Request Quote
Sale Sold out
Taxes included. Shipping calculated at checkout.

Description

The NVIDIA L40S is a powerful universal data center GPU built on Ada Lovelace architecture, delivering breakthrough performance for generative AI, large language model inference and training, 3D rendering, and video workloads. With 48GB GDDR6 ECC memory, 18,176 CUDA cores, 568 fourth-generation Tensor Cores with FP8 support, and 142 third-generation RT Cores, the L40S reaches up to 1,466 TFLOPS FP8 performance. Designed for enterprise 24/7 operations, it features passive cooling, PCIe 4.0 x16 interface, and NEBS Level 3 readiness in a dual-slot form factor.

Features

Fourth-generation Tensor Cores with FP8 precision for accelerated AI training and inference
- Third-generation RT Cores delivering up to 2x real-time ray-tracing performance
- Transformer Engine optimized for large language model acceleration
- NVIDIA DLSS 3 with AI-powered frame generation for enhanced graphics performance
- 48GB GDDR6 ECC memory for reliable operation and large model support
- Hardware support for structural sparsity and optimized TF32 format
- NVIDIA vGPU software support for virtualized graphics and compute workloads
- SR-IOV capability for hardware-based GPU partitioning
- Secure Boot with Root of Trust for enhanced data center security
- NEBS Level 3 readiness for telecommunications and enterprise deployments
- Passive cooling design for quiet operation and high-density installations
- Support for DirectX 12.07, Vulkan 1.18, OpenGL 4.68, and compute APIs
- Four DisplayPort 1.4a outputs for multi-monitor configurations
- PCIe 4.0 x16 interface with 64 GB/s bidirectional bandwidth
- Dual-slot form factor compatible with standard enterprise servers
- Optimized for 24/7 data center operations with enterprise-grade reliability

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU Architecture: NVIDIA Ada Lovelace
- Memory: 48GB GDDR6 with ECC
- Memory Bandwidth: 864 GB/s
- Memory Interface: 384-bit
- CUDA Cores: 18,176
- Tensor Cores: 568 (4th Generation) with FP8 support
- RT Cores: 142 (3rd Generation)
- Peak FP8 Performance: 1,466 TFLOPS (with sparsity)
- Peak FP32 Performance: 91.6 TFLOPS
- RT Core Performance: 212 TFLOPS
- System Interface: PCIe 4.0 x16
- Display Outputs: 4x DisplayPort 1.4a
- Form Factor: Dual-slot, full height, full length (10.5" L x 4.4" H)
- Cooling: Passive (fanless)
- Maximum Power Consumption: 350W
- API Support: OpenCL, DirectX 12.07, Vulkan 1.18, OpenGL 4.68, DirectCompute, OpenACC
- Virtualization: NVIDIA vGPU support, SR-IOV capable
- Security: Secure Boot with Root of Trust
- Certifications: NEBS Level 3 ready

FAQs

Q: What workloads is the NVIDIA L40S optimized for?
A: The L40S is optimized for multi-workload data center deployments including generative AI and large language model inference and training, 3D graphics rendering with real-time ray tracing, video transcoding and media acceleration, virtual desktop infrastructure, and multimodal AI applications combining audio, speech, 2D/3D, and video processing.

Q: How does the L40S compare to the previous generation A40 GPU?
A: The L40S delivers up to 5x higher inference performance than the A40, features fourth-generation Tensor Cores with FP8 precision support for faster AI training and inference, third-generation RT Cores for improved ray tracing, and DLSS 3 frame generation technology. Both cards feature 48GB memory, but the L40S is built on the newer Ada Lovelace architecture with significantly improved compute performance.

Q: Does the L40S support multi-GPU configurations?
A: The L40S uses PCIe 4.0 x16 for system connectivity and does not support NVLink interconnects or MIG partitioning. Multiple L40S cards can be installed in the same server for increased parallel processing capacity, but inter-GPU communication occurs over the PCIe bus rather than dedicated high-speed interconnects.

Q: What cooling and power requirements does the L40S have?
A: The L40S uses passive cooling and requires adequate server airflow for thermal management. It has a maximum power consumption of 350W and occupies a dual-slot, full-height, full-length PCIe form factor. The passive design enables quiet operation and high-density deployments in properly ventilated enterprise data center servers.

Q: Can the L40S be used for virtualization and multi-tenant environments?
A: Yes, the L40S supports NVIDIA vGPU software for graphics and compute virtualization, enabling multiple users to share GPU resources. It includes SR-IOV capabilities for hardware-based partitioning and features Secure Boot with Root of Trust for enhanced security in multi-tenant data center deployments.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us