Mistral Small 3.2 24B Instruct vs Gemma 3 27B it
Mistral Small 3.2 24B Instruct scores higher on the intelligence index, 7 against 5. Both take the same machine to start: the cheapest here that runs either is the Framework Desktop, 32GB, at $1,269. On it, they run at much the same speed: 9.7 and 9.8 tok/s, both estimated from memory bandwidth. At 500k tokens a day the Framework Desktop, 32GB pays for itself in 110 years running Gemma 3 27B it, against 163 years running Mistral Small 3.2 24B Instruct.
| Mistral Small 3.2 24B Instruct | Gemma 3 27B it | |
|---|---|---|
| Intelligence index | 7 | 5 |
| Class | Below every hosted tier | Below every hosted tier |
| Weights | 14 GB | 17 GB |
| Needs at 32k | 20 GB | 20 GB |
| Quantisation | Q4_K_M | Q4_K_M |
| Parameters | 24B | 27.4B |
| Max context | 128k | 128k |
| API price per 1M | $0.075 in / $0.2 out | $0.08 in / $0.45 out |
| Licence | Apache 2.0 | Gemma Terms of Use |
| Machines here that run it | 30 of 37 | 30 of 37 |
| Cheapest machine that runs it | Strix Halo Framework Desktop, 32GB $1,269 | Strix Halo Framework Desktop, 32GB $1,269 |
| Summarising | good | good |
| Translation | good | good |
| Everyday coding | usable | usable |
| Reasoning & maths | usable | usable |
| Agentic work | usable | don’t |
Run Mistral Small 3.2 24B Instruct on the Framework Desktop, 32GB · or Gemma 3 27B it
Ratings are coarse on purpose: they say what a model is usable for, not where it places to the decimal.
Side by side on the Framework Desktop, 32GB
The cheapest machine that runs either model is the same one, so this is the pair doing the same work on the same hardware: Strix Halo Framework Desktop, 32GB, at $1,269.
| Mistral Small 3.2 24B Instruct | Gemma 3 27B it | |
|---|---|---|
| Speed at 32k | 9.7 tok/s estimated | 9.8 tok/s estimated |
| Pay-back on this machine | Pays back in 163 years | Pays back in 110 years |
| API cost per month | $1.26 | $1.57 |
Run Mistral Small 3.2 24B Instruct on the Framework Desktop, 32GB · or Gemma 3 27B it
How much use it takes to pay for the machine
Everything above is at 500k tokens a day. Pay-back moves with how much you actually run, so here are both models at the five levels of use the calculator names, on the Framework Desktop, 32GB. Gemma 3 27B it pays for it sooner at every level of use, so which of them to run does not turn on how hard you work it.
| A day's use | Mistral Small 3.2 24B Instruct | Gemma 3 27B it |
|---|---|---|
| 50ka few chats a day | 1,633 years | 1,105 years |
| 200klight assistant use | 408 years | 276 years |
| 1Ma moderate coding-assistant day | 82 years | 55 years |
| 4Mheavy coding with an agent | 20 years | 14 years |
| 20Magents running most of the day | 6.1 yearsits ceiling | 4.1 yearsits ceiling |
The Framework Desktop, 32GB cannot generate 20M tokens a day on either model: at most 13.5M on Mistral Small 3.2 24B Instruct and 13.5M on Gemma 3 27B it. Both figures on that row are for the most it can do.
Run Mistral Small 3.2 24B Instruct at 20M tokens a day · or Gemma 3 27B it
Memory is not what separates them
Mistral Small 3.2 24B Instruct needs 20 GB of memory at 32k of context and Gemma 3 27B it needs 20 GB. Every machine priced here that runs one runs the other, at 32k of context. So the choice between them is what each is good at, how fast it runs and what the same work costs on an API, not what you have to buy to hold it.
Mistral Small 3.2 24B Instruct is also head to head with Qwen3 32B above it on the leaderboard, Qwen3 14B below it and Devstral Small 2 24B, the current Mistral nearest it in size. Gemma 3 27B it is also head to head with Qwen3 8B above it on the leaderboard, Ministral 3 8B below it and Gemma 4 31B it, the current Gemma nearest it in size. Two more models need much the same memory: Qwen3.6 27B and Devstral Small 2 24B.
The assumptions behind both columns
Both columns use the same usage: 500k tokens a day at 15:1 input to output, 32k of context, $0.17 per kWh, and today's API prices held flat. Speeds marked estimated are worked out from memory bandwidth rather than measured, and pay-back scales with them. Where nobody rents an open model by the token, its API prices are the nearest hosted model's, named beside them. Machines are the 37 here with a published price that are still sold. Change any of it in the calculator.
More head to head: every machine that runs Mistral Small 3.2 24B Instruct · every machine that runs Gemma 3 27B it · every other match-up · both against the frontier · the quickest pay-back at each level of use