NVIDIA GeForce RTX 5090, 32GB vs AMD Radeon AI PRO R9700, 32GB for local AI
Both hold 27 of the 39 open models here. The Radeon AI PRO R9700, 32GB costs $700 less, both priced as the card alone, without the PC around either. On Qwen3.8 27B, the strongest model both hold, the GeForce RTX 5090, 32GB is about 2.2× faster: 66 tok/s against 30, both estimated from memory bandwidth. The Radeon AI PRO R9700, 32GB pays for itself sooner, in 17 years against 25 years at 500k tokens a day.
| NVIDIA GeForce RTX 5090, 32GB | AMD Radeon AI PRO R9700, 32GB | |
|---|---|---|
| Price | $1,999card only | $1,299card only |
| Memory | 32 GB | 32 GB |
| Usable by the GPU | 31 GB | 31 GB |
| Memory bandwidth | 1792 GB/s | 640 GB/s |
| Power under load | 575 W | 300 W |
| Models that fit | 27 | 27 |
| Best model it runs | Qwen3.8 27B | Qwen3.8 27B |
| Speed on that model | 66 tok/s estimated | 30 tok/s estimated |
| Pay-back on that model | Pays back in 25 years | Pays back in 17 years |
Run the numbers on the NVIDIA GeForce RTX 5090, 32GB · or the AMD Radeon AI PRO R9700, 32GB
How much use it takes to pay back
Everything above is at 500k tokens a day. Pay-back moves with how much you actually run, so here are both machines on Qwen3.8 27B, the strongest model both hold, at the five levels of use the calculator names. The AMD Radeon AI PRO R9700, 32GB pays back sooner at every level of use, so this is not a choice that turns on how hard you work it.
| A day's use | NVIDIA GeForce RTX 5090, 32GB | AMD Radeon AI PRO R9700, 32GB |
|---|---|---|
| 50ka few chats a day | 254 years | 167 years |
| 200klight assistant use | 64 years | 42 years |
| 1Ma moderate coding-assistant day | 13 years | 8.3 years |
| 4Mheavy coding with an agent | 3.2 years | 2.1 years |
| 20Magents running most of the day | 7.6 months | 5.0 months |
Run the NVIDIA GeForce RTX 5090, 32GB at 20M tokens a day · or the AMD Radeon AI PRO R9700, 32GB
Memory is not what separates them
Every model on this list that fits one machine fits the other, and not only at 32k of context: across all 39 models the calculator counts, at every context from 4k to 256k, there is no model one holds and the other does not. Both leave 31 GB to the GPU. So the choice between them is speed, price and power, not what they can hold.
The assumptions behind both columns
Both columns use the same usage: 500k tokens a day at 15:1 input to output, 32k of context, $0.17 per kWh, and today's API prices held flat. Speeds marked estimated are worked out from memory bandwidth rather than measured, and pay-back scales with them. Graphics cards are priced as the card alone, so add the PC around one before comparing it with a complete computer. Change any of it in the calculator.
More head to head: everything the NVIDIA GeForce RTX 5090, 32GB runs · everything the AMD Radeon AI PRO R9700, 32GB runs · every other match-up · the quickest pay-back at each level of use · every model against the frontier