Compute Comparison
vs
All comparisons →

Volta

NVIDIA V100 32GB

wins

2

The V100 was NVIDIA's flagship AI GPU before the A100.

Legacy ML trainingBudget HPCMature framework support

Ampere

NVIDIA A10

wins

2

The NVIDIA A10 is a versatile Ampere GPU with 24GB GDDR6 and 150W TDP.

InferenceGraphics renderingVirtual workstations

Performance comparison

Visual bars — winner highlighted

NVIDIA V100 32GB

Metric

NVIDIA A10

FP16 TFLOPS

125 T
125 T

Bandwidth

900 GB/s
600 GB/s

VRAM

322 GB
246 GB

TDP (lower=better)

300 W
150 W

TFLOPS/watt

0.42
0.83

GB/s per watt

3.00
4.00

Full specification table

SpecNVIDIA V100 32GBNVIDIA A10
ArchitectureVoltaAmpere
VRAM32 GB HBM224 GB GDDR6
VRAM typeHBM2GDDR6
Memory bandwidth900 GB/s600 GB/s
FP16 TFLOPS125 TFLOPS125 TFLOPS
TDP300 W150 W
TFLOPS/watt0.42 T/W0.83 T/W
GB/s per watt3.00 GB/s/W4.00 GB/s/W
NVLinkYesNo
Release year20182021

Live cloud pricing

On-demand $/hr across providers — updated in real time

Loading prices…

Power efficiency analysis

TFLOPS/watt and GB/s/watt — critical for data center TCO

NVIDIA V100 32GB

FP16 TFLOPS125 TFLOPS
TDP300 W
TFLOPS/watt0.417 T/W
GB/s per watt3.00 GB/s/W

NVIDIA A10

FP16 TFLOPS125 TFLOPS
TDP150 W
TFLOPS/watt0.833 T/W
GB/s per watt4.00 GB/s/W

TFLOPS/watt measures compute efficiency — how much AI throughput you get per watt of power consumed. For data centers with PUE of 1.2–1.5, a 10% improvement in TFLOPS/watt translates directly to lower electricity costs and cooling requirements. GB/s/watt measures memory bandwidth efficiency, which is the binding constraint for memory-bound LLM inference workloads.

When to choose each GPU

Choose NVIDIA V100 32GB for:

  • Legacy ML training
  • Budget HPC
  • Mature framework support

Choose NVIDIA A10 for:

  • Inference
  • Graphics rendering
  • Virtual workstations