Skip to product information
1 of 1

Cisco

Cisco UCSX-GPU-H100-80 | H100 80GB PCIe GPU Accelerator, 350W, Dual-Slot FHFL

SKU:UCSX-GPU-H100-80

Stock Status: Enquire

Request Quote
Sale Sold out
Taxes included. Shipping calculated at checkout.

Description

The Cisco UCSX-GPU-H100-80 is an NVIDIA H100 Tensor Core GPU accelerator designed for UCS X-Series modular servers, delivering 80GB HBM2e memory and 350W TDP for enterprise AI training and high-performance computing workloads. Featuring fourth-generation Tensor Cores with FP8 precision support and a Transformer Engine, this dual-slot full-height full-length PCIe Gen5 card provides exceptional performance for large language models, deep learning inference, and data analytics. The H100 offers Multi-Instance GPU capability for workload isolation and supports NVLink for multi-GPU scaling in demanding AI deployments.

Features

Fourth-generation Tensor Cores with FP8, FP16, BFLOAT16, TF32, and FP64 precision support
- Transformer Engine for hardware-accelerated FP8 precision with automatic mixed-precision training
- 80GB HBM2e high-bandwidth memory with ECC for data integrity in mission-critical workloads
- 2TB/s memory bandwidth for high-throughput AI training and inference
- PCIe Gen5 x16 interface with 128GB/s bidirectional bandwidth
- Multi-Instance GPU (MIG) support for up to 7 isolated GPU instances with hardware-enforced partitioning
- NVLink 4.0 support with 600GB/s bandwidth for multi-GPU scaling and distributed training
- 3026 TFLOPS FP8 Tensor Core performance for accelerated deep learning workloads
- Hopper GH100 architecture built on TSMC 4nm process technology
- 7x NVDEC and 7x JPEG hardware decode engines for video processing workloads
- Full-Height Full-Length (FHFL) dual-slot form factor for compatibility with Cisco UCS X-Series PCIe nodes
- Passive cooling design optimized for data center server environments
- Support for CUDA, TensorRT, PyTorch, TensorFlow, and leading AI frameworks
- Hardware-accelerated encryption and secure boot for enterprise security requirements

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU: NVIDIA H100 Tensor Core
- GPU Memory: 80GB HBM2e with ECC
- Memory Bandwidth: 2TB/s (2000 GB/s)
- Compute Performance: 3026 TFLOPS FP8 Tensor Core (with sparsity), 1513 TFLOPS FP8 (dense), 51 TFLOPS FP32
- Interface: PCIe Gen5 x16 (128GB/s)
- Form Factor: Full-Height Full-Length (FHFL), Dual-Slot (2-slot width)
- Thermal Design Power (TDP): 350W
- Cooling: Passive (requires server airflow)
- Tensor Cores: 4th Generation with FP8, FP16, BFLOAT16, TF32, FP64 support
- Transformer Engine: Hardware-accelerated FP8 precision with automatic mixed-precision
- Multi-Instance GPU (MIG): Up to 7 isolated GPU instances
- NVLink Support: 600GB/s bandwidth for multi-GPU configurations
- Video Decode: 7x NVDEC, 7x JPEG decoders
- Compatible Platform: Cisco UCS X-Series (UCS X210c M7 compute nodes, UCS X440p PCIe nodes)

FAQs

Q: What UCS platforms support the UCSX-GPU-H100-80?
A: This GPU accelerator is designed for Cisco UCS X-Series modular systems, specifically the UCS X210c M7 compute nodes and the UCS X440p PCIe node installed in the UCS X9508 chassis. It requires PCIe Gen5 connectivity and proper riser configuration for dual-slot FHFL cards.

Q: What is the difference between the H100 PCIe and SXM variants?
A: The PCIe variant (UCSX-GPU-H100-80) uses HBM2e memory at 2TB/s bandwidth and has a 350W TDP, optimized for single-GPU inference and flexible server integration. The SXM variant uses HBM3 at 3.35TB/s with up to 700W TDP and higher NVLink bandwidth, designed for large-scale multi-GPU training clusters.

Q: Does this GPU support Multi-Instance GPU (MIG) partitioning?
A: Yes, the H100 supports MIG technology, allowing a single GPU to be partitioned into up to 7 isolated instances. Each MIG instance receives dedicated CUDA cores, Tensor Cores, L2 cache, and HBM memory with hardware-enforced isolation, ideal for multi-tenant AI inference workloads.

Q: What AI workloads benefit most from the H100's Transformer Engine?
A: The Transformer Engine with FP8 precision is optimized for large language models and transformer-based architectures, delivering up to 3-4x faster training compared to previous generation GPUs on models like GPT-3, BERT, and Llama. It automatically balances FP8 and FP16 precision to maximize throughput while maintaining model accuracy.

Q: What power and cooling requirements does this GPU have?
A: The UCSX-GPU-H100-80 has a 350W TDP and uses passive cooling, relying on the UCS chassis airflow for thermal management. The UCS X-Series chassis must be configured with adequate power supplies and cooling capacity to support the GPU's power draw and heat dissipation requirements.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us