AMD · CDNA 5 · 432 GB HBM4 · EAM · 2026
0 providers · 0 configurations · prices as of Sep 29, 2026, 10:15 UTC
Daily Instinct MI455X prices per GPU-hour, one point per day across all tracked providers.
Tracking since : the chart fills in with one point per day.
No provider currently lists the Instinct MI455X. Prices refresh every 2 hours.
Memory for the weights plus one 32K-token sequence of context, how many GPUs hold it, and what that replica costs at the cheapest in-stock on-demand rate.
| Model | Params | Memory | GPUs | $/hr | $/mo |
|---|---|---|---|---|---|
gpt-oss-20b OpenAI · 20.9B | 20.9B · 3.6B active | 20 GiB | 1 | — | — |
Gemma 3 27B Google · 27.4B | 27.4B | 28 GiB | 1 | — | — |
Qwen3-32B Alibaba · 32.8B | 32.8B | 39 GiB | 1 | — | — |
Llama 3.3 70B Instruct Meta · 70.6B | 70.6B | 76 GiB | 1 | — | — |
gpt-oss-120b OpenAI · 117B | 117B · 5.1B active | 110 GiB | 1 | — | — |
Qwen3-235B-A22B Alibaba · 235B | 235B · 22B active | 225 GiB | 1 | — | — |
DeepSeek-R1 DeepSeek · 684B | 684B · 37B active | 640 GiB | 2 | — | — |
Kimi K3 Moonshot AI · 2.78T | 2.78T · 104B active | 2.53 TiB | 8 | — | — |
Assumes 90% of each GPU's memory is usable and rounds up to 1, 2, 4, 8… GPUs for tensor parallelism (up to 64). $/hr is the GPU count times the cheapest in-stock on-demand rate per GPU; check that the provider rents instances that large. Serving many requests at once needs more memory for context.
GPUs with similar compute and memory, and their lowest on-demand price right now.
432 GB HBM4 with 23,300 GB/s of memory bandwidth, CDNA 5 architecture (2026), EAM form factor. Dense peak: 5.03 PF FP16 / BF16, 20.1 PF FP8, 40.3 PF FP4, 5.03 PF INT8, 0.315 PF FP32, 0.005 PF FP64.
Prices were collected on Sep 29, 2026, 10:15 UTC from provider APIs and public catalogs and refresh about every 2 hours. They are indicative list prices; always confirm on the provider's website before booking.