Skip to product information
1 of 1

Cisco

Cisco UCSX-GPUA100-80-D | NVIDIA A100 80GB HBM2e GPU, 300W, PCIe Gen4, MIG

SKU:UCSX-GPUA100-80-D

Stock Status: Enquire

Request Quote
Sale Sold out
Taxes included. Shipping calculated at checkout.

Description

The Cisco UCSX-GPUA100-80-D is an NVIDIA A100 Tensor Core GPU designed for enterprise AI training, HPC, and data analytics workloads in Cisco UCS X-Series modular systems. Featuring 80GB of HBM2e memory with up to 1.94 TB/s bandwidth, 300W TDP, and PCIe Gen4 x16 connectivity, this passively cooled graphics card delivers exceptional compute performance for demanding data center applications. Multi-Instance GPU technology enables partitioning into up to seven isolated instances, maximizing resource utilization for multi-tenant AI infrastructure.

Features

NVIDIA Ampere architecture with 6,912 CUDA cores and 3rd-generation Tensor Cores
- 80GB HBM2e high-bandwidth memory with up to 1.94 TB/s bandwidth for large-scale AI models
- Multi-Instance GPU (MIG) partitioning into up to 7 isolated instances for multi-tenant workloads
- Structural sparsity support for up to 2x inference performance on sparse models
- TF32 precision delivering up to 20x higher AI training performance vs previous generation
- Mixed-precision support: FP64, FP32, TF32, FP16, BFLOAT16, INT8, INT4
- PCIe Gen4 x16 interface with 64 GB/s bidirectional bandwidth
- NVLink Bridge support for multi-GPU scaling up to 600 GB/s interconnect bandwidth
- SR-IOV virtualization support with up to 20 virtual functions (VF)
- Passive cooling design optimized for high-density data center rack deployments
- ECC memory protection for mission-critical HPC and enterprise AI applications
- Power cable included for simplified installation
- Designed and validated for Cisco UCS X-Series modular infrastructure

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU: NVIDIA A100 Tensor Core (Ampere architecture)
- Memory: 80GB HBM2e
- Memory Bandwidth: Up to 1.94 TB/s (1,935 GB/s)
- Interface: PCIe Gen4 x16
- Form Factor: Dual-slot, full-height, full-length (FHFL)
- Cooling: Passive heatsink (requires system airflow)
- TDP: 300W
- Power Connector: 8-pin CPU auxiliary power connector
- Multi-Instance GPU: Up to 7 MIG instances (10GB each)
- Peak FP64: 9.7 TFLOPS (19.5 TFLOPS with Tensor Cores)
- Peak FP32: 19.5 TFLOPS
- Peak TF32 Tensor Core: 156 TFLOPS (312 TFLOPS with sparsity)
- Peak FP16/BFLOAT16 Tensor Core: 312 TFLOPS (624 TFLOPS with sparsity)
- Peak INT8 Tensor Core: 624 TOPS (1,248 TOPS with sparsity)
- NVLink Support: Yes (via optional NVLink Bridge for dual-GPU configurations)
- Compatibility: Cisco UCS X440p PCIe Node (Riser 1A, Riser 2A Gen4 slots)
- Included: Power cable
- Weight: Approximately 1,170g (board only)

FAQs

Q: What servers is this GPU compatible with?
A: The UCSX-GPUA100-80-D is designed for Cisco UCS X-Series modular systems, specifically the UCS X440p PCIe Node with M7 servers. It installs in Riser 1A or Riser 2A (Gen4) slots and requires PCIe Gen4 x16 support. Each server can support up to two of these GPUs depending on configuration.

Q: What is Multi-Instance GPU (MIG) and how does it work?
A: MIG technology allows the A100 80GB to be partitioned into up to seven independent GPU instances, each with dedicated compute, memory (10GB per instance), and cache resources. This enables multiple users or workloads to share a single GPU simultaneously while maintaining hardware-level isolation, maximizing utilization in multi-tenant AI and inference environments.

Q: What cooling requirements does this GPU have?
A: This is a passive (fanless) GPU that relies entirely on system airflow for cooling. It requires deployment in a data center server chassis with adequate forced-air cooling to operate within its 300W TDP thermal envelope. It is not suitable for standard workstations or systems with insufficient airflow.

Q: Can I connect multiple A100 GPUs together?
A: Yes. Up to two A100 80GB PCIe GPUs can be connected using three NVIDIA NVLink Bridges (sold separately) to deliver 600 GB/s bi-directional bandwidth between GPUs—10x faster than PCIe Gen4 alone. This configuration is ideal for large-scale AI training and HPC workloads requiring multi-GPU parallelism.

Q: What workloads is the A100 80GB optimized for?
A: The A100 80GB excels at AI model training and inference, deep learning, natural language processing, recommendation systems, HPC simulations, computational fluid dynamics, molecular dynamics, and scientific computing. Its large 80GB memory capacity handles massive models and datasets, while Tensor Core acceleration dramatically speeds FP16, TF32, and mixed-precision workloads.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us