GLM-4.7-Flash vs Gemma 4 12B
GLM-4.7-Flash scores higher on the intelligence index — 15 against 14. Gemma 4 12B is the smaller download at 7.1 GB, so it runs on cheaper hardware.
| GLM-4.7-Flash | Gemma 4 12B | |
|---|---|---|
| Intelligence index | 15 | 14 |
| Class | Haiku-class | Below every hosted tier |
| Weights | 18 GB | 7.1 GB |
| Quantisation | Q4_K_M | Q4_K_M |
| Parameters | 31.2B (3B active) | 12B |
| Max context | 198k | 256k |
| API price per 1M | $0.061 in / $0.4 out | $0.08 in / $0.13 out |
| Licence | MIT | Apache 2.0 |
| Cheapest machine that runs it | Strix Halo Framework Desktop, 32GB $1,269 | Mac mini M6, 16GB $899 |
| Summarising | not rated | not rated |
| Translation | not rated | not rated |
| Everyday coding | not rated | not rated |
| Reasoning & maths | not rated | not rated |
| Agentic work | not rated | not rated |
Ratings are coarse on purpose. Speeds and pay-back depend on the machine — open either model's page for the full list, or see both against the frontier. Context is 32k throughout.