Overview
NVIDIA A10 Tensor Core GPU delivers 24GB GDDR6 and 250 TOPS INT8 in a passive single-slot PCIe card designed for rack-server inference. At 31.2 TFLOPS FP32 and 125 FP16 TFLOPS with 600 GB/s bandwidth, it's positioned for cost-efficient enterprise AI inference where multiple A10 cards per server provide horizontal scaling. Widely deployed for LLM serving and real-time computer vision.
Key Features
- 24GB GDDR6 — 600 GB/s bandwidth
- 31.2 TFLOPS FP32 / 250 TOPS INT8 for inference
- Passive blower cooling — server rack compatible
- PCIe 4.0 x16, 150W TDP
- NVIDIA Ampere Tensor Cores with INT8 optimization
Ideal For
Enterprises deploying multi-GPU inference servers where cost-per-inference matters — the A10's passive cooling and INT8 acceleration make it efficient for production LLM serving and computer vision at scale.
No specifications available.
$2,470.00
Prices may vary. Verify on vendor site.
Tags
llm-inferencecomputer-visionmodel-training
