Skip to product information
1 of 1

Cisco

Cisco HCI-GPU-A30-M6 | A30 GPU, 24GB HBM2, PCIe 4.0, Passive Cooling, 180W

SKU:HCI-GPU-A30-M6

Stock Status: Enquire

Request Quote
Sale Sold out
Shipping calculated at checkout.

Description

The Cisco HCI-GPU-A30-M6 is an NVIDIA A30 Tensor Core GPU designed for Cisco hyperconverged infrastructure servers. Featuring 24GB of HBM2 memory with 933GB/s bandwidth and 180W TDP in a passive-cooled, full-height, full-length dual-slot form factor, this GPU accelerates AI inference, machine learning training, high-performance computing, and virtualized workloads with Multi-Instance GPU support. Built on NVIDIA Ampere architecture with PCIe 4.0 x16 interface, it delivers versatile compute acceleration for mainstream enterprise data center applications.

Features

NVIDIA Ampere architecture with third-generation Tensor Cores for AI acceleration
- 24GB HBM2 memory with ECC for reliability in enterprise environments
- 933GB/s memory bandwidth for fast data transfer and large dataset processing
- Multi-Instance GPU (MIG) technology partitions GPU into up to 4 isolated instances
- Supports diverse precision formats: FP64, FP32, TF32, FP16, INT8, INT4
- PCIe 4.0 x16 interface for high-speed host communication
- Structural sparsity acceleration for optimized neural network inference
- NVLink Bridge support enables 2x A30 GPU interconnect with 200GB/s bidirectional bandwidth
- NVIDIA vGPU software support for virtual desktop infrastructure and RDSH deployments
- Secure Boot and firmware protection features for enterprise security
- Compatible with NVIDIA AI Enterprise, Virtual Compute Server, and NGC catalog
- Passive cooling design optimized for Cisco server thermal management
- Dual-slot form factor with full-height, full-length PCIe card dimensions
- Dynamic workload switching between AI inference, training, and HPC applications

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU: NVIDIA A30 Tensor Core GPU, Ampere architecture
- Memory: 24GB HBM2 with ECC, 933GB/s bandwidth
- CUDA Cores: 3,584
- Tensor Cores: Third-generation with TF32, FP16, FP64 support
- Interface: PCIe 4.0 x16
- Form Factor: Full-height, full-length (FHFL), dual-slot (double-wide)
- Thermal Design Power (TDP): 180W
- Cooling: Passive (requires server airflow management)
- Multi-Instance GPU (MIG): Up to 4 isolated GPU instances
- Performance: 10.3 TFLOPS FP32, 5.2 TFLOPS FP64, 165 TFLOPS INT8
- NVLink Support: NVLink Bridge compatible for dual-GPU configurations
- Compatible Servers: Cisco HCI 240C M6, C240 M6, C245 M6 series (specific riser slots)
- Server Integration: Requires Cisco-specific SBIOS ID for compatibility

FAQs

Q: Which Cisco servers support the HCI-GPU-A30-M6 GPU?
A: The HCI-GPU-A30-M6 is designed for Cisco hyperconverged infrastructure servers including the HCI 240C M6 and UCS C240 M6 series. It can be installed in riser slots 2, 5, or 7 depending on server configuration. As a double-wide GPU, it occupies two adjacent PCIe slots. The GPU requires Cisco-specific SBIOS firmware and must be procured from Cisco for proper server integration.

Q: What is Multi-Instance GPU (MIG) and how does it benefit enterprise deployments?
A: The A30 GPU supports MIG technology, allowing the GPU to be partitioned into up to four fully isolated GPU instances, each with dedicated high-bandwidth memory, cache, and compute cores. This enables IT administrators to offer right-sized GPU acceleration for different workloads simultaneously, maximizing utilization and providing secure resource allocation for AI inference, training, and HPC applications running concurrently.

Q: What workloads are best suited for the NVIDIA A30 GPU?
A: The A30 is optimized for mainstream enterprise AI inference at scale, machine learning training, high-performance computing with FP64 precision, data analytics, and virtual desktop infrastructure with NVIDIA vGPU software. Its versatile compute capabilities support diverse workloads including conversational AI, computer vision, scientific simulation, and dynamic workload balancing between production inference and model retraining.

Q: What are the power and cooling requirements for this GPU?
A: The HCI-GPU-A30-M6 has a 180W TDP and uses passive cooling, requiring proper server airflow management. Cisco servers with GPUs require low-profile heatsinks and a dedicated GPU air duct accessory. Organizations should use the Cisco UCS power calculator to ensure adequate PSU capacity when configuring servers with multiple GPUs or high-performance components.

Q: Can I mix the HCI-GPU-A30-M6 with other GPU models in the same server?
A: No, Cisco servers do not support mixing different GPU models or brands within the same system. All GPUs in a server must be identical. If you plan to install multiple GPUs, they must all be HCI-GPU-A30-M6 cards. You can install up to three A30 GPUs in compatible Cisco HCI servers depending on riser configuration and power supply capacity.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us