Gemma 4 31B it vs Nemotron 3.5 Lightning 30B-A3B
Gemma 4 31B it scores higher on the intelligence index, 15 against 14. Both take the same machine to start: the cheapest here that runs either is the Radeon AI PRO R9700, 32GB, at $1,299. The Radeon AI PRO R9700, 32GB is priced as the card alone, without the PC around it. On it, Nemotron 3.5 Lightning 30B-A3B is about 5.8× quicker: 127 tok/s against 22, both estimated from memory bandwidth. At 500k tokens a day the Radeon AI PRO R9700, 32GB pays for itself in 88 years running Nemotron 3.5 Lightning 30B-A3B, against 110 years running Gemma 4 31B it.
| Gemma 4 31B it | Nemotron 3.5 Lightning 30B-A3B | |
|---|---|---|
| Intelligence index | 15 | 14 |
| Class | Haiku-class | Below every hosted tier |
| Weights | 20 GB | 25 GB |
| Needs at 32k | 26 GB | 26 GB |
| Quantisation | Q4_K_M | Q4_K_M |
| Parameters | 31.3B | 31.6B (3B active) |
| Max context | 256k | 256k |
| API price per 1M | $0.09 in / $0.34 out | $0.08 in / $0.2 out |
| Licence | Apache 2.0 | OpenMDW 1.1 |
| Machines here that run it | 27 of 37 | 27 of 37 |
| Cheapest machine that runs it | AMD Radeon AI PRO R9700, 32GB $1,299card only | AMD Radeon AI PRO R9700, 32GB $1,299card only |
| Summarising | not rated | not rated |
| Translation | not rated | not rated |
| Everyday coding | not rated | not rated |
| Reasoning & maths | not rated | not rated |
| Agentic work | not rated | not rated |
Run Gemma 4 31B it on the Radeon AI PRO R9700, 32GB · or Nemotron 3.5 Lightning 30B-A3B
Ratings are coarse on purpose: they say what a model is usable for, not where it places to the decimal.
Side by side on the Radeon AI PRO R9700, 32GB
The cheapest machine that runs either model is the same one, so this is the pair doing the same work on the same hardware: AMD Radeon AI PRO R9700, 32GB, at $1,299card only.
| Gemma 4 31B it | Nemotron 3.5 Lightning 30B-A3B | |
|---|---|---|
| Speed at 32k | 22 tok/s estimated | 127 tok/s estimated |
| Pay-back on this machine | Pays back in 110 years | Pays back in 88 years |
| API cost per month | $1.61 | $1.33 |
Run Gemma 4 31B it on the Radeon AI PRO R9700, 32GB · or Nemotron 3.5 Lightning 30B-A3B
How much use it takes to pay for the machine
Everything above is at 500k tokens a day. Pay-back moves with how much you actually run, so here are both models at the five levels of use the calculator names, on the Radeon AI PRO R9700, 32GB. Nemotron 3.5 Lightning 30B-A3B pays for it sooner at every level of use, so which of them to run does not turn on how hard you work it.
| A day's use | Gemma 4 31B it | Nemotron 3.5 Lightning 30B-A3B |
|---|---|---|
| 50ka few chats a day | 1,101 years | 883 years |
| 200klight assistant use | 275 years | 221 years |
| 1Ma moderate coding-assistant day | 55 years | 44 years |
| 4Mheavy coding with an agent | 14 years | 11 years |
| 20Magents running most of the day | 2.8 years | 2.2 years |
Run Gemma 4 31B it at 20M tokens a day · or Nemotron 3.5 Lightning 30B-A3B
Memory is not what separates them
Gemma 4 31B it needs 26 GB of memory at 32k of context and Nemotron 3.5 Lightning 30B-A3B needs 26 GB. Every machine priced here that runs one runs the other, at 32k of context. So the choice between them is what each is good at, how fast it runs and what the same work costs on an API, not what you have to buy to hold it.
Gemma 4 31B it is also head to head with Qwen3.5 122B-A10B above it on the leaderboard, Granite 4.2 30B below it and Gemma 3 27B it, the last-generation Gemma nearest it in size. Nemotron 3.5 Lightning 30B-A3B is also head to head with Qwen3.5 9B above it on the leaderboard and Qwen3 235B-A22B Instruct 2507 below it.
The assumptions behind both columns
Both columns use the same usage: 500k tokens a day at 15:1 input to output, 32k of context, $0.17 per kWh, and today's API prices held flat. Speeds marked estimated are worked out from memory bandwidth rather than measured, and pay-back scales with them. Where nobody rents an open model by the token, its API prices are the nearest hosted model's, named beside them. Machines are the 37 here with a published price that are still sold. Change any of it in the calculator.
More head to head: every machine that runs Gemma 4 31B it · every machine that runs Nemotron 3.5 Lightning 30B-A3B · every other match-up · both against the frontier · the quickest pay-back at each level of use