Skip to product information
1 of 1

Cisco

Cisco HX-GPU-A30 | A30 24GB HBM2e Tensor Core GPU, PCIe 4.0, 180W Passive

SKU:HX-GPU-A30

Stock Status: Enquire

Request Quote
Sale Sold out
Shipping calculated at checkout.

Description

The Cisco HX-GPU-A30 is an NVIDIA A30 Tensor Core GPU designed for AI inference, machine learning, and high-performance computing workloads in Cisco HyperFlex hyperconverged infrastructure. Built on NVIDIA Ampere architecture with 24GB HBM2e memory and 933GB/s bandwidth, it delivers exceptional performance for AI training and inference, scientific simulations, and data analytics. The passive-cooled, PCIe 4.0 form factor provides energy-efficient acceleration for mainstream enterprise servers, with Multi-Instance GPU technology enabling up to four isolated GPU instances for optimal resource utilization.

Features

NVIDIA Ampere architecture with 3rd generation Tensor Cores for AI and HPC acceleration
- 24GB HBM2e high-bandwidth memory (933GB/s) for demanding data-intensive workloads
- Multi-Instance GPU (MIG) technology partitions GPU into up to 4 isolated instances
- TF32 precision delivers 10X AI performance boost over previous generation with no code changes
- Structured sparsity support provides up to 2X performance improvement for sparse models
- FP64 Tensor Cores accelerate scientific computing and HPC applications
- PCIe 4.0 x16 interface for maximum bandwidth to host system
- NVLink support enables multi-GPU scaling for larger workloads
- vGPU support for virtual desktop infrastructure and virtualized AI workloads
- NVIDIA AI Enterprise certified for VMware vSphere environments
- Passive cooling design for reliable, quiet operation in enterprise servers
- 180W TDP for energy-efficient AI acceleration in mainstream data centers
- Compatible with NVIDIA CUDA, TensorRT, Magnum IO SDK, and NGC containers
- Optimized for real-time inference with INT8 precision support

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU: NVIDIA A30 Tensor Core (Ampere Architecture)
- Memory: 24GB HBM2e
- Memory Bandwidth: 933GB/s
- Interface: PCIe 4.0 x16
- CUDA Cores: 3584 stream processors
- Tensor Cores: 224 (3rd generation)
- Peak FP32 Performance: 10.32 TFLOPS
- Peak TF32 Tensor Performance: 165 TFLOPS (with sparsity)
- Peak FP64 Performance: 5.2 TFLOPS
- Peak FP64 Tensor Core Performance: 10.3 TFLOPS
- Thermal Design Power (TDP): 180W
- Cooling: Passive heatsink
- Form Factor: Dual-slot, full-height, PCIe card
- Multi-Instance GPU (MIG): Up to 4 GPU instances
- NVLink Support: Yes (up to 2 GPUs with NVLink Bridge)
- Virtualization: vGPU ready, supports NVIDIA AI Enterprise
- Compatibility: Cisco HyperFlex HX240/HX245 M6 platforms

FAQs

Q: What workloads is the HX-GPU-A30 optimized for?
A: The A30 GPU is designed for diverse enterprise AI workloads including scalable AI inference, machine learning training, high-performance computing applications, and data analytics. Its balanced compute and memory configuration makes it ideal for mainstream server deployments requiring AI acceleration without extreme power demands.

Q: What is Multi-Instance GPU (MIG) and how does it benefit deployment?
A: MIG technology allows the A30 to be partitioned into up to four fully isolated GPU instances, each with dedicated memory, cache, and compute cores. This enables multiple workloads or users to securely share a single GPU with guaranteed quality of service, maximizing utilization and providing right-sized acceleration for various job sizes.

Q: Which Cisco platforms support the HX-GPU-A30?
A: The HX-GPU-A30 is supported on Cisco HyperFlex 240 and 245 M6 platform series running HX Data Platform Release 5.0(2b) or later with Qualified FI/Server Firmware 4.2(2d). It is designed specifically for integration with Cisco HyperFlex hyperconverged infrastructure.

Q: Why is the A30 passively cooled at 180W?
A: The passive cooling design allows the A30 to operate quietly and reliably in standard enterprise server environments with chassis-based airflow. The 180W TDP provides a balance between compute performance and energy efficiency, making it suitable for mainstream data center deployment without requiring active cooling solutions.

Q: What precision formats does the A30 support?
A: The A30 Tensor Cores support a wide range of precision formats including FP64, FP32, TF32 (Tensor Float 32), BFLOAT16, FP16, and INT8, enabling acceleration across diverse AI training, inference, and HPC workloads. TF32 provides up to 10X faster performance than FP32 with zero code changes for AI applications.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us