Skip to product information
1 of 1

NVIDIA

NVIDIA 900-2G179-0020-101 | A2 Tensor Core 16GB GPU, PCIe 4.0 x8, 40-60W TDP

SKU:900-2G179-0020-101

Stock Status: Enquire

$2,385.79 inc. GST
$2,168.90 ex. GST
Sale Sold out
Taxes included. Shipping calculated at checkout.

Description

The NVIDIA A2 Tensor Core GPU delivers entry-level AI inference acceleration with exceptional power efficiency for edge deployments and compact server environments. Built on the Ampere architecture with 16GB GDDR6 memory and 1280 CUDA cores, it provides 4.5 TFLOPS FP32 performance and up to 72 TOPS INT4 tensor performance in a single-slot, low-profile form factor. Ideal for intelligent video analytics, machine learning inference, and virtual desktop infrastructure workloads where space and power constraints are critical.

Features

NVIDIA Ampere architecture with 3rd generation Tensor Cores for AI inference acceleration
- Single-slot, low-profile form factor with passive cooling for flexible server integration
- Configurable 40-60W TDP for optimized power efficiency in edge deployments
- 16GB GDDR6 memory with 200 GB/s bandwidth for large AI models and datasets
- Support for FP32, TF32, FP16, BFLOAT16, INT8, and INT4 precision formats
- NVIDIA sparsity technology delivers up to 2x performance improvement for compatible models
- 10 RT Cores (2nd generation) for ray tracing acceleration
- Dual video decoders with AV1 decode support for intelligent video analytics
- Single video encoder for efficient video processing workflows
- PCIe 4.0 x8 interface for high-speed data transfer and reduced CPU bottlenecks
- NVIDIA vGPU software support for virtual desktop infrastructure and cloud AI workloads
- Secure boot capabilities with trusted code authentication and firmware rollback protection
- Compatibility with NVIDIA AI Enterprise, Triton Inference Server, and CUDA-X libraries
- No external power connector required - powered entirely through PCIe slot
- RoHS compliant for environmentally responsible deployments

Warranty

All products sold by XS Network Tech include a 12-month warranty on both new and used items. Our in-house technical team thoroughly tests used hardware prior to sale to ensure enterprise-grade reliability.

All technical data should be verified on the manufacturer data sheets.

View full details

specs-tabs

Collapsible content

Technical Specifications

FAQs

Technical Specifications

GPU Architecture: NVIDIA Ampere
- CUDA Cores: 1280
- Tensor Cores: 40 (3rd generation)
- RT Cores: 10 (2nd generation)
- Memory: 16GB GDDR6
- Memory Interface: 128-bit
- Memory Bandwidth: 200 GB/s
- Base Clock: 1440 MHz
- Boost Clock: 1770 MHz
- Peak FP32 Performance: 4.5 TFLOPS
- Peak FP16 Tensor Core: 18 TFLOPS (36 TFLOPS with sparsity)
- Peak INT8 Tensor Core: 36 TOPS (72 TOPS with sparsity)
- Peak INT4 Tensor Core: 72 TOPS (144 TOPS with sparsity)
- Interface: PCIe 4.0 x8
- Form Factor: Single-slot, low-profile, full-height bracket included
- Cooling: Passive (fanless)
- TDP: 40-60W (configurable)
- Video Engines: 1x encoder, 2x decoders (includes AV1 decode)
- vGPU Support: NVIDIA Virtual PC, Virtual Applications, RTX Virtual Workstation, AI Enterprise, Virtual Compute Server
- Dimensions: 16.8 cm (depth) x 6.88 cm (height)
- Compliance: RoHS

FAQs

Q: What AI workloads is the NVIDIA A2 optimized for?
A: The A2 is designed for entry-level AI inference workloads including intelligent video analytics, edge AI deployments, machine learning inference, and virtual desktop infrastructure. It excels in scenarios requiring low power consumption and compact form factors while delivering up to 20x better inference performance compared to CPU-only servers.

Q: What server form factors support this GPU?
A: The A2 supports both low-profile and full-height server configurations thanks to its single-slot design and included brackets. The passive cooling and 40-60W configurable TDP make it compatible with standard 1U and 2U rackmount servers, tower servers, and edge computing platforms with PCIe 4.0 x8 slots.

Q: Does this GPU support virtualization?
A: Yes, the A2 supports NVIDIA vGPU software including Virtual PC, Virtual Applications, RTX Virtual Workstation, AI Enterprise, and Virtual Compute Server. This enables multiple users to share GPU resources for virtual desktop infrastructure and cloud-based AI inference applications.

Q: What precision formats does the A2 Tensor Core support?
A: The A2's 3rd generation Tensor Cores support multiple precision formats including FP32, TF32, FP16, BFLOAT16, INT8, and INT4. It also supports NVIDIA's automatic mixed precision capabilities and sparsity acceleration, which can double effective throughput for compatible AI models.

Q: Can multiple A2 GPUs be installed in the same server?
A: Yes, the A2's single-slot, low-profile design and passive cooling enable dense multi-GPU configurations in standard servers. Multiple A2 GPUs can be deployed to scale inference capacity or run parallel workloads, making it ideal for edge data centers and consolidated AI deployments.

Recently Viewed

  • Request a Quote

    Looking for competitive pricing? Submit a request, and our team will provide a tailored quote that fits your needs.

  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Call us Now  
  • Contact Us Directly

    Have a question or need immediate assistance? Call us for expert advice and real-time support.

    Contact us