…
Cloud rental price, compute per dollar and specs, side by side.
Open in the interactive comparer →| Spec / price | L40S | A100 SXM 80GB |
|---|---|---|
| Memory | 48 GB GDDR6 | 80 GB HBM2ebetter |
| Memory bandwidth | 864 GB/s | 2,039 GB/sbetter |
| FP16 / BF16 (dense) | 0.362 PFbetter | 0.312 PF |
| FP8 (dense) | 0.733 PF | — |
| FP4 (dense) | — | — |
| INT8 (dense) | 0.733 PFbetter | 0.624 PF |
| TF32 (dense) | 0.183 PFbetter | 0.156 PF |
| FP32 (dense) | 0.0916 PFbetter | 0.0195 PF |
| FP64 (dense) | 0.0014 PF | 0.0195 PFbetter |
| Power | 350 W | 400 W |
| Architecture | Ada Lovelace (2023) | Ampere (2020) |
| Lowest on-demand | $0.60better | $0.81 |
| Median on-demand | $1.50better | $2.35 |
| Lowest spot | $0.38better | $0.90 |
| FP16/BF16 $ per PFLOP·h | $1.66better | $2.60 |
| Providers | 17 | 17 |
The L40S has 1.2× more dense FP16/BF16 compute than the A100 SXM 80GB. The A100 SXM 80GB has 2.4× more memory bandwidth than the L40S. Real-world speedups depend on the workload: training and prefill scale with compute, while LLM decoding scales mostly with memory bandwidth.
Right now the L40S is cheaper per unit of compute: $1.66 vs $2.60 per PFLOP-hour at the lowest on-demand prices.
Running one GPU around the clock at the lowest on-demand price costs about $440 per month for the L40S and $591 for the A100 SXM 80GB.
The A100 SXM 80GB has more memory: 80 GB vs 48 GB.