Ampere
NVIDIA RTX A6000
wins
3
The RTX A6000 is NVIDIA's Ampere professional workstation GPU with 48GB GDDR6 and NVLink support.
Ampere
NVIDIA A40
wins
0
The NVIDIA A40 offers 48GB GDDR6 at a lower price point than the A100.
Performance comparison
Visual bars — winner highlighted
NVIDIA RTX A6000
Metric
NVIDIA A40
FP16 TFLOPS
Bandwidth
VRAM
TDP (lower=better)
TFLOPS/watt
GB/s per watt
Full specification table
| Spec | NVIDIA RTX A6000 | NVIDIA A40 |
|---|---|---|
| Architecture | Ampere | Ampere |
| VRAM | 48 GB GDDR6 | 48 GB GDDR6 |
| VRAM type | GDDR6 | GDDR6 |
| Memory bandwidth | 768 GB/s | 696 GB/s |
| FP16 TFLOPS | 154 TFLOPS | 150 TFLOPS |
| TDP | 300 W | 300 W |
| TFLOPS/watt | 0.51 T/W | 0.50 T/W |
| GB/s per watt | 2.56 GB/s/W | 2.32 GB/s/W |
| NVLink | Yes | No |
| Release year | 2020 | 2020 |
Live cloud pricing
On-demand $/hr across providers — updated in real time
Power efficiency analysis
TFLOPS/watt and GB/s/watt — critical for data center TCO
NVIDIA RTX A6000
NVIDIA A40
TFLOPS/watt measures compute efficiency — how much AI throughput you get per watt of power consumed. For data centers with PUE of 1.2–1.5, a 10% improvement in TFLOPS/watt translates directly to lower electricity costs and cooling requirements. GB/s/watt measures memory bandwidth efficiency, which is the binding constraint for memory-bound LLM inference workloads.
When to choose each GPU
Choose NVIDIA RTX A6000 for:
- Professional workstation AI
- Large model inference
- Rendering
Choose NVIDIA A40 for:
- Large model inference
- Professional rendering
- Multi-tenant AI
Popular GPU comparisons