Skip to product information
1 of 1

Cisco

Cisco UCSX-GPU-A100-80 | A100 Tensor Core GPU, 80GB HBM2e, 300W Passive, PCIe Gen4

SKU:UCSX-GPU-A100-80

Stock Status: Enquire

Request Quote
Sale Sold out
Taxes included. Shipping calculated at checkout.

Description

The Cisco UCSX-GPU-A100-80 is an NVIDIA A100 Tensor Core GPU with 80GB HBM2e memory, designed for AI training, inference, and high-performance computing workloads in Cisco UCS X-Series and C-Series servers. Featuring 6,912 CUDA cores, 432 third-generation Tensor Cores, and 1,935 GB/s memory bandwidth, this passive-cooled, 300W PCIe Gen4 x16 accelerator delivers exceptional performance for data science, machine learning, and scientific computing applications. The card includes a power cable and is optimized for double-wide PCIe riser configurations.

Features

NVIDIA Ampere architecture with 54 billion transistors on 7nm process
- 80GB HBM2e high-bandwidth memory with ECC for data integrity
- Third-generation Tensor Cores with TF32, BF16, and FP64 acceleration
- Multi-Instance GPU (MIG) support: partition into up to 7 isolated instances
- Structured sparsity for 2x AI inference and training acceleration
- PCIe Gen4 x16 interface with 64 GB/s bidirectional bandwidth
- Passive cooling design for high-density server deployments
- 312 TFLOPS FP16/BF16 peak performance (624 TFLOPS with sparsity)
- 1,935 GB/s memory bandwidth for large model and dataset support
- Double-wide form factor optimized for UCS X-Series and C-Series
- Unified memory architecture for efficient data transfer
- Native support for CUDA, TensorRT, RAPIDS, and leading AI frameworks
- SR-IOV support with up to 20 virtual functions for virtualized environments
- Compatible with NVIDIA AI Enterprise and NGC software stack

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU: NVIDIA Tesla A100 Tensor Core, Ampere architecture
- Memory: 80GB HBM2e, 5,120-bit memory interface
- Memory Bandwidth: 1,935 GB/s (PCIe variant)
- CUDA Cores: 6,912
- Tensor Cores: 432 (3rd generation)
- FP64 Performance: 9.7 TFLOPS (CUDA), 19.5 TFLOPS (Tensor Core)
- TF32 Performance: 156 TFLOPS (312 TFLOPS with sparsity)
- FP16/BF16 Performance: 312 TFLOPS (624 TFLOPS with sparsity)
- Interface: PCIe Gen4 x16
- TDP: 300W passive cooling
- Form Factor: Dual-slot, full-height
- Multi-Instance GPU: Up to 7 isolated GPU instances @ 10GB each
- ECC Memory: Enabled by default
- Compatible Platforms: Cisco UCS X-Series (UCSX-440P, X9508), UCS C-Series (C240 M6/M7)
- Riser Support: PCIe Gen4 x16 slots (Riser 1A, 2A, 3C)
- Included: Power cable

FAQs

Q: What Cisco UCS servers support the UCSX-GPU-A100-80?
A: This GPU is compatible with Cisco UCS X-Series modular systems (UCSX-440P PCIe node with X9508 chassis) and select UCS C-Series rack servers (C240 M6/M7). It requires PCIe Gen4 x16 slots (Riser 1A slot 2, Riser 2A slot 5, or Riser 3C slot 7). Each compatible server can support up to three of these GPUs, depending on riser configuration and power capacity.

Q: What is Multi-Instance GPU (MIG) and how does it work on the A100 80GB?
A: MIG allows the A100 to be partitioned into up to seven isolated GPU instances, each with 10GB of memory in the 80GB variant. Each instance operates with dedicated memory, compute resources, and quality-of-service guarantees, enabling multiple users or workloads to share a single GPU securely and efficiently. This is ideal for consolidating diverse AI inference, training, and HPC workloads in elastic data center environments.

Q: How does the PCIe A100 80GB differ from the SXM variant?
A: The PCIe variant (UCSX-GPU-A100-80) offers 300W TDP and 1,935 GB/s memory bandwidth, using a standard PCIe Gen4 x16 interface for broad server compatibility. The SXM variant runs at 400W TDP with 2,039 GB/s bandwidth and supports NVLink for multi-GPU interconnects at 600 GB/s. The PCIe model is ideal for standard rack servers, while SXM is optimized for NVIDIA HGX baseboard deployments.

Q: What precision formats does the A100 Tensor Core support?
A: The A100's third-generation Tensor Cores support FP64, FP32, TF32 (Tensor Float 32), BF16 (BFloat16), FP16, INT8, and INT4 precisions. TF32 delivers 20x the performance of FP32 on V100 for AI training without code changes, while structured sparsity can double throughput for FP16/BF16 and INT8 workloads, reaching up to 624 TFLOPS and 1,248 TOPS respectively.

Q: Why must A100 GPUs be procured from Cisco for UCS systems?
A: Cisco UCS systems require GPUs with a unique SBIOS ID for proper recognition by CIMC (Cisco Integrated Management Controller) and UCSM (UCS Manager). Only Cisco-validated A100 SKUs like the UCSX-GPU-A100-80 include the necessary firmware and management integration for full compatibility, monitoring, and inventory reporting within the UCS ecosystem.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us