Sunk Cost sunkcost.ai Data checked 2026-09-03

GMKtec EVO-X2: 64GB vs 128GB for local AI

The GMKtec EVO-X2, 128GB holds 33 of the 39 open models here, and the GMKtec EVO-X2, 64GB holds 27. The GMKtec EVO-X2, 64GB costs $1,300 less. On Qwen3.8 27B, the strongest model both hold, they run at much the same speed: 10 and 10 tok/s, both estimated from memory bandwidth. The GMKtec EVO-X2, 64GB pays for itself sooner, in 29 years against 46 years at 500k tokens a day.

Strix Halo GMKtec EVO-X2, 64GBStrix Halo GMKtec EVO-X2, 128GB
Price$2,200$3,500
Memory64 GB128 GB
Usable by the GPU48 GB96 GB
Memory bandwidth256 GB/s256 GB/s
Power under load150 W150 W
Models that fit2733
Best model it runsQwen3.8 27BQwen3.8 27B
Speed on that model10 tok/s estimated10 tok/s estimated
Pay-back on that modelPays back in 29 yearsPays back in 46 years

Run the numbers on the Strix Halo GMKtec EVO-X2, 64GB · or the Strix Halo GMKtec EVO-X2, 128GB

How much use it takes to pay back

Everything above is at 500k tokens a day. Pay-back moves with how much you actually run, so here are both machines on Qwen3.8 27B, the strongest model both hold, at the five levels of use the calculator names. The Strix Halo GMKtec EVO-X2, 64GB pays back sooner at every level of use, so this is not a choice that turns on how hard you work it.

A day's useStrix Halo GMKtec EVO-X2, 64GBStrix Halo GMKtec EVO-X2, 128GB
50ka few chats a day291 years464 years
200klight assistant use73 years116 years
1Ma moderate coding-assistant day15 years23 years
4Mheavy coding with an agent3.6 years5.8 years
20Magents running most of the day12 monthsits ceiling20 monthsits ceiling

On Qwen3.8 27B neither machine can generate 20M tokens a day: the Strix Halo GMKtec EVO-X2, 64GB manages at most 14.3M and the Strix Halo GMKtec EVO-X2, 128GB at most 14.3M. Both figures on that row are for the most each can do.

Run the Strix Halo GMKtec EVO-X2, 64GB at 20M tokens a day · or the Strix Halo GMKtec EVO-X2, 128GB

What the extra memory buys

The Strix Halo GMKtec EVO-X2, 128GB holds 6 models the Strix Halo GMKtec EVO-X2, 64GB cannot at 32k of context. The strongest of them are what the difference in memory actually buys.

ModelWeightsNeeds at 32kOn the Strix Halo GMKtec EVO-X2, 128GB
Ling 3.0 flashHaiku-class78 GB82 GB20 tok/s estimated
Qwen3.5 122B-A10BHaiku-class78 GB79 GB20 tok/s estimated
gpt-oss-120bBelow every hosted tier63 GB65 GB37 tok/s measured
Mistral Small 4 (119B-2603)Below every hosted tier74 GB75 GB32 tok/s estimated
Qwen3-Coder NextBelow every hosted tier48 GB49 GB54 tok/s estimated
Devstral 2 123BBelow every hosted tier75 GB87 GB2.2 tok/s estimated

The assumptions behind both columns

Both columns use the same usage: 500k tokens a day at 15:1 input to output, 32k of context, $0.17 per kWh, and today's API prices held flat. Speeds marked estimated are worked out from memory bandwidth rather than measured, and pay-back scales with them. Graphics cards are priced as the card alone, so add the PC around one before comparing it with a complete computer. Change any of it in the calculator.

More head to head: everything the Strix Halo GMKtec EVO-X2, 64GB runs · everything the Strix Halo GMKtec EVO-X2, 128GB runs · every other match-up · the quickest pay-back at each level of use · every model against the frontier