Skip to product information
1 of 1

Cisco

Cisco HCI-GPU-A40 | A40 GPU, 48GB GDDR6 ECC, PCIe 4.0, Passive 300W

SKU:HCI-GPU-A40

Stock Status: Enquire

Request Quote
Sale Sold out
Shipping calculated at checkout.

 Trade supply only. Server & GPU quote requests need a company name, ABN and work email address. We do not quote Gmail, Outlook or Hotmail addresses, or supply for home use. 

20-min response

Need a quote or bulk pricing? Answered within 20 minutes, Sydney business hours.

Custom Built Servers

Need a different spec, quantity or configuration?

We build configured-to-order refurbished servers from Cisco, Dell, HPE, NVIDIA Lenovo and Supermicro. Tell us what you need and we'll come back with options within 24 hours.

Configure your server now →

Description

The Cisco HCI-GPU-A40 is a high-performance data center GPU accelerator built on NVIDIA Ampere architecture. Featuring 48GB GDDR6 ECC memory, 10,752 CUDA cores, and second-generation RT Cores, it delivers powerful acceleration for AI inference, 3D rendering, virtual workstations, and compute-intensive workloads. The passive cooling design is optimized for rack server deployments with adequate airflow.

Features

NVIDIA Ampere architecture with 10,752 CUDA cores for parallel processing
- 48GB GDDR6 ECC memory for large dataset handling and data integrity
- Second-generation RT Cores for real-time ray tracing acceleration
- Third-generation Tensor Cores for AI and deep learning workloads
- PCIe 4.0 x16 interface with 64 GB/s bidirectional bandwidth
- NVLink connectivity for multi-GPU configurations up to 96GB combined memory
- NVIDIA vGPU support for virtual desktop infrastructure (VDI) deployments
- SR-IOV capability for GPU virtualization and resource sharing
- 696 GB/s memory bandwidth for data-intensive applications
- Passive thermal solution optimized for high-density server environments
- Support for mixed-precision computing (FP32, FP16, TF32, INT8, INT4)
- Hardware video encode/decode engines
- Headless operation for data center and cloud deployments
- Suitable for AI inference, rendering farms, virtual workstations, and HPC

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU: NVIDIA A40 based on Ampere architecture
- GPU Memory: 48GB GDDR6 with ECC
- Memory Bandwidth: 696 GB/s
- Memory Interface: 384-bit
- CUDA Cores: 10,752
- RT Cores: 84 (second-generation)
- Tensor Cores: 336 (third-generation)
- FP32 Performance: 37.4 TFLOPS
- RT Core Performance: 73.1 TFLOPS
- Tensor Performance: Up to 299.4 TFLOPS (FP16)
- Host Interface: PCIe 4.0 x16
- NVLink Support: Yes, 112.5 GB/s bidirectional
- Thermal Design Power (TDP): 300W
- Power Connector: 8-pin PCIe
- Cooling: Passive heatsink (requires chassis airflow)
- Form Factor: Full-height, full-length (FHFL), dual-slot
- Display Outputs: Headless configuration (no external connectors)
- Virtualization: NVIDIA vGPU support (vPC, vApp, RTX vWS, vCS)
- APIs: CUDA, DirectX 12, OpenGL, Vulkan

FAQs

Q: What workloads is the A40 GPU optimized for?
A: The A40 is designed for data center visual computing workloads including AI inference and training, 3D rendering, ray tracing, virtual workstations (VDI), simulation, CAD/CAE applications, and high-performance computing tasks requiring large memory capacity.

Q: Does this GPU have video display outputs?
A: The Cisco HCI-GPU-A40 is configured as a headless data center card with no external display connectors. It is intended for server integration and virtualization deployments rather than direct display connection.

Q: Can two A40 GPUs be linked together?
A: Yes, two A40 GPUs can be connected via NVLink bridge to scale memory capacity from 48GB to 96GB and increase GPU-to-GPU interconnect bandwidth, enabling unified memory for larger datasets and improved performance in supported applications.

Q: What are the cooling requirements for this GPU?
A: The A40 uses passive cooling and requires adequate chassis airflow to dissipate the 300W thermal load. It is designed for rack-mount servers with optimized front-to-back or side-to-side airflow configurations.

Q: Is this GPU compatible with NVIDIA vGPU software?
A: Yes, the A40 supports NVIDIA vGPU software including vPC, vApps, RTX Virtual Workstation, and Virtual Compute Server, enabling virtualized GPU resources for multiple concurrent users in VDI and cloud environments.

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us