Skip to product information
1 of 1

Cisco

Cisco HCI-GPU-L40 | NVIDIA L40 GPU, 48GB GDDR6, 300W, PCIe Gen4 x16

SKU:HCI-GPU-L40

Stock Status: Enquire

Request Quote
Sale Sold out
Shipping calculated at checkout.

Description

The Cisco HCI-GPU-L40 delivers enterprise-grade GPU acceleration powered by the NVIDIA L40 with Ada Lovelace architecture. Featuring 48GB of GDDR6 memory with ECC, 300W TDP, and PCIe Gen4 x16 interface, this 2-slot full-height full-length GPU is optimized for AI inference, virtual workstations, 3D rendering, and graphics-intensive data center workloads. Designed for deployment in Cisco UCS X-Series and C-Series servers, it provides the computational power required for generative AI, machine learning training, and professional visualization applications.

Features

NVIDIA Ada Lovelace architecture with 18,176 CUDA cores for parallel processing
- 48GB GDDR6 ECC memory for large AI models and datasets
- 4th generation Tensor Cores for accelerated AI training and inference
- 3rd generation RT Cores for hardware-accelerated ray tracing
- PCIe Gen4 x16 interface with 64 GB/s bidirectional bandwidth
- Support for mixed-precision computing (FP32, TF32, FP16, INT8)
- NVIDIA vGPU software support for virtualized GPU workloads
- 4x DisplayPort 1.4 outputs for multi-monitor configurations
- ECC memory protection for enterprise reliability
- Passive cooling design for server integration
- Full-height full-length (FHFL) 2-slot form factor
- Advanced shading and simulation capabilities
- Hardware-accelerated video encoding and decoding
- NVIDIA CUDA, cuDNN, and TensorRT software compatibility
- Optimized for generative AI, LLM inference, and graphics workloads

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU: NVIDIA L40 based on Ada Lovelace architecture
- Memory: 48GB GDDR6 with ECC
- Memory Bandwidth: 864 GB/s
- Interface: PCIe Gen4 x16 (64 GB/s bidirectional)
- CUDA Cores: 18,176
- Tensor Cores: 568 (4th Generation)
- RT Cores: 142 (3rd Generation)
- FP32 Performance: 90.5 TFLOPS
- RT Core Performance: 209 TFLOPS
- Maximum Power: 300W TDP
- Form Factor: 2-slot full-height full-length (FHFL)
- Cooling: Passive
- Display Outputs: 4x DisplayPort connectors
- Process Technology: 4nm
- Transistor Count: 76.3 billion
- Compatible with: Cisco UCS X-Series (UCSX) and C-Series (UCSC) servers

FAQs

Q: What workloads is the NVIDIA L40 GPU optimized for?
A: The NVIDIA L40 is optimized for AI inference, generative AI applications, virtual workstation deployments (with NVIDIA vWS software), 3D rendering, professional visualization, video processing, and graphics-intensive data center workloads. It excels at neural graphics, mixed-precision compute tasks, and ray tracing applications.

Q: Which Cisco server platforms support the HCI-GPU-L40?
A: This GPU is designed for Cisco UCS X-Series modular systems (particularly with the UCSX-440P PCIe Node in the X9508 chassis) and Cisco UCS C-Series rack servers. It requires PCIe Gen4 x16 slots and appropriate power delivery. Verify server compatibility and cooling requirements before deployment.

Q: What are the power and cooling requirements for the L40?
A: The NVIDIA L40 has a 300W TDP and uses passive cooling, meaning it relies on server airflow rather than onboard fans. Ensure your server has adequate forced-air cooling and a compatible power supply with PCIe power connectors. The card occupies two PCIe slots due to its dual-slot form factor.

Q: Can the L40 be used for virtual desktop infrastructure (VDI)?
A: Yes, when combined with NVIDIA RTX Virtual Workstation (vWS) software, the L40 can be virtualized to deliver high-performance GPU-accelerated workstation instances to remote users. With 48GB of memory, it supports multiple concurrent virtual workstation sessions for design, engineering, and compute-intensive applications.

Q: How does the L40 compare to the L40S?
A: The L40S is the successor model with enhanced performance. It features nearly double the TF32 and FP16 Tensor Core performance and increases maximum power from 300W to 350W. The standard L40 remains a cost-effective option for workloads that don't require the additional compute density of the L40S.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us