Skip to product information
1 of 1

Cisco

Cisco HCI-GPU-A40-M6 | A40 GPU, 48GB GDDR6 ECC, PCIe 4.0 x16, Passive Cooling

SKU:HCI-GPU-A40-M6

Stock Status: Enquire

Request Quote
Sale Sold out
Shipping calculated at checkout.

Description

The Cisco HCI-GPU-A40-M6 is an NVIDIA A40 data center GPU built on Ampere architecture, designed for AI inference, rendering, virtualization, and HPC workloads. With 10,752 CUDA cores, 336 third-generation Tensor Cores, 84 second-generation RT Cores, and 48GB of ECC GDDR6 memory, it delivers exceptional performance for enterprise compute, visual computing, and AI applications. The passive cooling design requires adequate chassis airflow and is optimized for high-density rack server deployments.

Features

NVIDIA Ampere architecture with 10,752 CUDA cores for parallel processing
- 48GB GDDR6 ECC memory for large dataset handling and error correction
- 336 third-generation Tensor Cores for AI training and inference acceleration
- 84 second-generation RT Cores for real-time ray tracing and rendering
- PCIe 4.0 x16 interface with 31.5 GB/s bidirectional bandwidth
- NVIDIA NVLink support for multi-GPU configurations up to 96GB total memory
- Passive thermal design optimized for high-density rack server deployments
- NVIDIA vGPU software support for virtual desktop infrastructure and remote workstations
- Tensor Float 32 (TF32) precision for accelerated AI model training
- Hardware support for structural sparsity to double inference throughput
- DLSS, AI denoising, and enhanced editing capabilities for graphics workloads
- Secure and measured boot with hardware root of trust
- Compatible with Cisco UCS C-Series M6 and HyperFlex M6 servers
- Requires unique SBIOS ID for CIMC and UCSM integration
- Backward compatible with PCIe Gen 3 systems

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU Model: NVIDIA A40 with Ampere architecture
- GPU Memory: 48GB GDDR6 with ECC
- CUDA Cores: 10,752
- Tensor Cores: 336 (3rd generation)
- RT Cores: 84 (2nd generation)
- Memory Interface: 384-bit
- Memory Bandwidth: 696 GB/s
- Bus Interface: PCIe 4.0 x16 (backward compatible with PCIe Gen 3)
- Interconnect: NVIDIA NVLink 112.5 GB/s bidirectional
- Thermal Design Power: 300W
- Cooling: Passive heatsink (requires chassis airflow)
- Form Factor: Full-height, full-length (FHFL), dual-slot
- Display Outputs: Headless (no external display connectors by default)
- Peak FP32 Performance: 37.4 TFLOPS
- Virtualization Support: NVIDIA vGPU, vWS, Virtual Compute Server

FAQs

Q: What workloads is the NVIDIA A40 GPU optimized for?
A: The A40 is optimized for AI inference and training, real-time ray tracing, professional visualization, virtual workstations, rendering farms, scientific simulation, data science, and HPC workloads. It combines compute power with advanced graphics capabilities.

Q: Does this GPU support virtualization?
A: Yes, the A40 supports NVIDIA vGPU software for virtual workstation delivery and VDI environments, allowing multiple users to share GPU resources remotely. It is compatible with NVIDIA RTX Virtual Workstation and Virtual Compute Server software.

Q: Is the GPU headless or does it have display outputs?
A: The A40 is configured as headless by default with physical display connectors disabled for data center operation. Display port functionality can be enabled via Cisco management tools if needed for specific use cases.

Q: What cooling requirements does this GPU have?
A: The HCI-GPU-A40-M6 uses passive cooling with a heatsink and requires adequate chassis airflow to dissipate 300W TDP. It is designed for installation in Cisco UCS rack servers with proper ventilation.

Q: Can I scale memory capacity beyond 48GB?
A: Yes, you can connect two A40 GPUs with NVLink to scale GPU memory from 48GB to 96GB, increasing bandwidth and enabling larger dataset processing for compute-intensive applications.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us