CDNA 4
AMD Instinct MI355X
wins
4
The AMD Instinct MI355X is AMD's CDNA 4 flagship.
CDNA 3
AMD Instinct MI300X
wins
0
The AMD MI300X leads all GPUs on memory capacity (192GB HBM3) and bandwidth (5.3 TB/s), making it exceptional for serving very large language models that don't fit on H100.
Performance comparison
Visual bars — winner highlighted
AMD Instinct MI355X
Metric
AMD Instinct MI300X
FP16 TFLOPS
Bandwidth
VRAM
TDP (lower=better)
TFLOPS/watt
GB/s per watt
Full specification table
| Spec | AMD Instinct MI355X | AMD Instinct MI300X |
|---|---|---|
| Architecture | CDNA 4 | CDNA 3 |
| VRAM | 288 GB HBM3e | 192 GB HBM3 |
| VRAM type | HBM3e | HBM3 |
| Memory bandwidth | 8,000 GB/s | 5,300 GB/s |
| FP16 TFLOPS | 2,615 TFLOPS | 1,307 TFLOPS |
| TDP | 750 W | 750 W |
| TFLOPS/watt | 3.49 T/W | 1.74 T/W |
| GB/s per watt | 10.67 GB/s/W | 7.07 GB/s/W |
| NVLink | No | No |
| Release year | 2025 | 2023 |
Live cloud pricing
On-demand $/hr across providers — updated in real time
Power efficiency analysis
TFLOPS/watt and GB/s/watt — critical for data center TCO
AMD Instinct MI355X
AMD Instinct MI300X
TFLOPS/watt measures compute efficiency — how much AI throughput you get per watt of power consumed. For data centers with PUE of 1.2–1.5, a 10% improvement in TFLOPS/watt translates directly to lower electricity costs and cooling requirements. GB/s/watt measures memory bandwidth efficiency, which is the binding constraint for memory-bound LLM inference workloads.
When to choose each GPU
Choose AMD Instinct MI355X for:
- Large LLM inference
- Agent workloads
- Memory-bound AI
Choose AMD Instinct MI300X for:
- Memory-bound LLM inference
- Large model serving
- HPC
Popular GPU comparisons