Llama 3.3 70B Instruct vs LFM2.5 2.6B
Neither has a clear lead on the index. LFM2.5 2.6B is the smaller download at 1.7 GB, so it runs on cheaper hardware.
| Llama 3.3 70B Instruct | LFM2.5 2.6B | |
|---|---|---|
| Intelligence index | 8 | 8 |
| Class | Below every hosted tier | Below every hosted tier |
| Weights | 43 GB | 1.7 GB |
| Quantisation | Q4_K_M | Q4_K_M |
| Parameters | 70.6B | 2.7B |
| Max context | 128k | 128k |
| API price per 1M | $0.1 in / $0.32 out | $0 in / $0 out |
| Licence | Llama 3.3 Community License | LFM 1.0 (see repo) |
| Cheapest machine that runs it | Strix Halo Framework Desktop, 128GB $3,449 | Mac mini M6, 16GB $899 |
| Summarising | good | not rated |
| Translation | good | not rated |
| Everyday coding | good | not rated |
| Reasoning & maths | usable | not rated |
| Agentic work | usable | not rated |
Ratings are coarse on purpose. Speeds and pay-back depend on the machine — open either model's page for the full list, or see both against the frontier. Context is 32k throughout.