Local LLMs that fit in 64 GB
24 of 47 current models have a listed weight option that needs at most 85% of 64 GB.
against memory
OpenRouter evals, GPQA Diamond · data from 2026-10-10
Fewer than 8 models with a score fit this memory budget, so there is no chart. The values are in the table.
| # | Model | Configuration | Min memory | GPQA Diamond | Frontier |
|---|---|---|---|---|---|
| 1 | Qwen3.8 27B | 27B Original weights | 54.3 GB ◇ | 81.9% | On frontier |
| 2 | Nemotron 3.5 Lightning | 30B-A3B Original weights BF16 | 62.3 GB ◇ | 69.8% | — |
Scores were measured on hosted endpoints, so each model is plotted once, at the memory need of its original weights. A quantized file needs less memory and may score differently. Lines connect the frontier points and do not predict results in between. Source: OpenRouter evals (openrouter.ai) via OpenRouter (openrouter.ai/rankings). openrouter.ai
Models
Models with a GPQA Diamond score come first, highest score first. The rest are newest first.
An option is listed when its memory need is at most 85% of 64 GB. The figure after each option is an estimate from file size, an 8K-token context and 0.5 GB for the runtime, unless the publisher states one.