Qwen3.5 122B-A10B vs GLM-4.5-Air
Qwen3.5 122B-A10B scores higher on the intelligence index, 16 against 11. Both take the same machine to start: the cheapest here that runs either is the Framework Desktop, 128GB, at $3,449. On it, Qwen3.5 122B-A10B is about 2× quicker: 20 tok/s against 10, both estimated from memory bandwidth. At 500k tokens a day the Framework Desktop, 128GB pays for itself in 53 years running Qwen3.5 122B-A10B, against 139 years running GLM-4.5-Air.
| Qwen3.5 122B-A10B | GLM-4.5-Air | |
|---|---|---|
| Intelligence index | 16 | 11 |
| Class | Haiku-class | Below every hosted tier |
| Weights | 78 GB | 73 GB |
| Needs at 32k | 79 GB | 79 GB |
| Quantisation | UD-Q4_K_M | Q4_K_M |
| Parameters | 125.1B (10B active) | 110.5B (12B active) |
| Max context | 256k | 128k |
| API price per 1M | $0.26 in / $2.08 out | $0.13 in / $0.85 out |
| Licence | Apache 2.0 | MIT |
| Machines here that run it | 12 of 37 | 12 of 37 |
| Cheapest machine that runs it | Strix Halo Framework Desktop, 128GB $3,449 | Strix Halo Framework Desktop, 128GB $3,449 |
| Summarising | not rated | good |
| Translation | not rated | good |
| Everyday coding | not rated | good |
| Reasoning & maths | not rated | usable |
| Agentic work | not rated | good |
Run Qwen3.5 122B-A10B on the Framework Desktop, 128GB · or GLM-4.5-Air
Ratings are coarse on purpose: they say what a model is usable for, not where it places to the decimal.
Side by side on the Framework Desktop, 128GB
The cheapest machine that runs either model is the same one, so this is the pair doing the same work on the same hardware: Strix Halo Framework Desktop, 128GB, at $3,449.
| Qwen3.5 122B-A10B | GLM-4.5-Air | |
|---|---|---|
| Speed at 32k | 20 tok/s estimated | 10 tok/s estimated |
| Pay-back on this machine | Pays back in 53 years | Pays back in 139 years |
| API cost per month | $5.69 | $2.66 |
Run Qwen3.5 122B-A10B on the Framework Desktop, 128GB · or GLM-4.5-Air
How much use it takes to pay for the machine
Everything above is at 500k tokens a day. Pay-back moves with how much you actually run, so here are both models at the five levels of use the calculator names, on the Framework Desktop, 128GB. Qwen3.5 122B-A10B pays for it sooner at every level of use, so which of them to run does not turn on how hard you work it.
| A day's use | Qwen3.5 122B-A10B | GLM-4.5-Air |
|---|---|---|
| 50ka few chats a day | 533 years | 1,392 years |
| 200klight assistant use | 133 years | 348 years |
| 1Ma moderate coding-assistant day | 27 years | 70 years |
| 4Mheavy coding with an agent | 6.7 years | 17 years |
| 20Magents running most of the day | 16 months | 5.0 yearsits ceiling |
The Framework Desktop, 128GB generates at most 13.8M tokens a day on GLM-4.5-Air, so that column's figure at 20M tokens a day is for the most it can do, not for the whole of what was asked.
Run Qwen3.5 122B-A10B at 20M tokens a day · or GLM-4.5-Air
Memory is not what separates them
Qwen3.5 122B-A10B needs 79 GB of memory at 32k of context and GLM-4.5-Air needs 79 GB. Every machine priced here that runs one runs the other, at 32k of context. So the choice between them is what each is good at, how fast it runs and what the same work costs on an API, not what you have to buy to hold it.
Qwen3.5 122B-A10B is also head to head with Gemma 4 26B-A4B above it on the leaderboard and Gemma 4 31B it below it. One more model needs much the same memory: Mistral Small 4 (119B-2603). GLM-4.5-Air is also head to head with Granite 4.2 8B above it on the leaderboard and Mistral Small 4 (119B-2603) below it. One more model needs much the same memory: Ling 3.0 flash.
The assumptions behind both columns
Both columns use the same usage: 500k tokens a day at 15:1 input to output, 32k of context, $0.17 per kWh, and today's API prices held flat. Speeds marked estimated are worked out from memory bandwidth rather than measured, and pay-back scales with them. Where nobody rents an open model by the token, its API prices are the nearest hosted model's, named beside them. Machines are the 37 here with a published price that are still sold. Change any of it in the calculator.
More head to head: every machine that runs Qwen3.5 122B-A10B · every machine that runs GLM-4.5-Air · every other match-up · both against the frontier · the quickest pay-back at each level of use