Skip to product information
1 of 1

Cisco

Cisco HX-GPU-A100-80 | A100 80GB HBM2e Graphics Card, 6912 CUDA, 300W

SKU:HX-GPU-A100-80

Stock Status: Enquire

Request Quote
Sale Sold out
Shipping calculated at checkout.

 Trade supply only. Server & GPU quote requests need a company name, ABN and work email address. We do not quote Gmail, Outlook or Hotmail addresses, or supply for home use. 

20-min response

Need a quote or bulk pricing? Answered within 20 minutes, Sydney business hours.

Custom Built Servers

Need a different spec, quantity or configuration?

We build configured-to-order refurbished servers from Cisco, Dell, HPE, NVIDIA Lenovo and Supermicro. Tell us what you need and we'll come back with options within 24 hours.

Configure your server now →

Description

The Cisco HX-GPU-A100-80 is a PCIe Gen4 compute accelerator based on the NVIDIA A100 Tensor Core GPU with 80GB HBM2e memory and 300W TDP. Designed for AI training, HPC workloads, and large-scale data analytics in Cisco UCS server environments, it features passive cooling and double-width form factor. This GPU delivers exceptional performance for deep learning, scientific computing, and inference tasks requiring massive memory capacity.

Features

NVIDIA Ampere architecture with 54.2 billion transistors on 7nm process
- Third-generation Tensor Cores with support for FP64, TF32, BF16, FP16, INT8, and INT4 precision
- Structural sparsity acceleration for 2x Tensor Core throughput
- 80GB HBM2e memory with ECC for data integrity
- Multi-Instance GPU (MIG) technology for workload isolation and GPU virtualization
- PCIe Gen4 x16 interface with 64 GB/s bidirectional bandwidth
- NVLink Bridge support for dual-GPU configurations with 600 GB/s interconnect
- Passive thermal solution optimized for high-airflow server environments
- Support for CUDA, cuDNN, TensorRT, RAPIDS, and major AI frameworks
- Double-width form factor with dual-slot mounting
- Unified memory architecture for efficient data access
- Asynchronous copy and compute engine overlap
- Hardware-accelerated ray tracing capabilities
- Designed for 24/7 data center operation
- Requires CIMC/UCSM integration with Cisco-specific SBIOS ID

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU Architecture: NVIDIA Ampere (GA100)
- CUDA Cores: 6,912
- Tensor Cores: 432 (3rd Generation)
- GPU Memory: 80GB HBM2e with ECC
- Memory Bandwidth: 1,935 GB/s (PCIe variant)
- Memory Interface: 5,120-bit
- Base Clock: 1,065 MHz / Boost Clock: 1,410 MHz
- FP64 Performance: 9.7 TFLOPS
- FP32 Performance: 19.5 TFLOPS
- TF32 Tensor Performance: 156 TFLOPS (312 TFLOPS with sparsity)
- FP16/BF16 Tensor Performance: 312 TFLOPS (624 TFLOPS with sparsity)
- INT8 Tensor Performance: 624 TOPS (1,248 TOPS with sparsity)
- Multi-Instance GPU (MIG): Up to 7 instances @ 10GB each
- Interface: PCIe Gen4 x16
- TDP: 300W
- Form Factor: Dual-slot, passive cooling, double-width
- Compatible with Cisco HX240c and select UCS servers
- Requires x16 PCIe support in Riser 1A slot 2, Riser 2A slot 5, or Riser 3C slot 7
- Up to 3 GPUs per supported server

FAQs

Q: What is Multi-Instance GPU (MIG) and how does it work on the A100?
A: MIG technology allows you to partition a single A100 GPU into up to seven isolated instances, each with dedicated memory, cache, and compute resources. This enables multiple workloads to run simultaneously with hardware-level isolation, improving GPU utilization and cost efficiency in multi-tenant or mixed workload environments.

Q: Which Cisco UCS servers are compatible with the HX-GPU-A100-80?
A: This GPU is compatible with Cisco HX240c and select UCS C-Series and X-Series servers that support PCIe Gen4 x16 expansion. It can be populated in Riser 1A slot 2, Riser 2A slot 5, or Riser 3C slot 7. Each compatible server can support up to three of these GPUs. The GPU requires a unique SBIOS ID registered with CIMC and UCSM for proper operation.

Q: What workloads benefit most from the 80GB memory configuration?
A: The 80GB HBM2e memory is ideal for training large language models, processing massive datasets in data analytics, running complex HPC simulations, and performing inference on memory-intensive AI models. The high memory capacity eliminates bottlenecks when working with models that exceed the 40GB available on smaller variants.

Q: Does this GPU support NVLink for multi-GPU configurations?
A: The HX-GPU-A100-80 PCIe variant supports NVLink Bridge for connecting two GPUs with 600 GB/s bidirectional bandwidth. For larger multi-GPU configurations beyond two units, GPUs communicate via PCIe Gen4 at 64 GB/s per GPU. The SXM form factor variants support full NVLink mesh topologies.

Q: What cooling requirements does this passive GPU have?
A: This is a passive cooled, 300W TDP GPU requiring robust chassis airflow for proper thermal management. It must be installed in a Cisco server with adequate front-to-back airflow and appropriate power supply capacity. The double-width form factor occupies two expansion slots.

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us