SpanveroHow it worksFind a modelCompare modelsPricing

The best open LLMs you can run on 8 GB of VRAM

Every open LLM in our catalog whose weights plus KV-cache actually fit in 8 GB of VRAM at its default quant — the size of an RTX 3060/4060 or an 8 GB laptop GPU. Ranked by popularity and recognition within the set that fits (the most-run models first), with the honest $0-on-your-own-hardware cost for each. You pick the one whose quality you like; we just guarantee it fits.

How this is ranked: Pure objective filter: 'best' = 'fits your 8 GB card.' VRAM is computed by our shared cost engine from params, quant and context — not a quality opinion. Within the fit set we order by popularity/recognition (real Hugging Face downloads — a recognized shortlist) and never claim a #1 is 'smartest.' Quality judgment is the user's; the 8 GB VRAM hub lists every fitting model, largest first.

Showing the top 40 of 231. See all →

Want every model that fits? All 231 models that run on 8 GB of VRAM →

More: all "best" lists · Outcome Lab · all models

Open the free Spanvero advisor → · Honest, $0-markup. © 2026 Cynosure LLC.

The weekly price index

A short email of real AI price moves, straight from the daily log — no hype. We're collecting the list now; the first issue goes out when it opens. Unsubscribe with one click.