SpanveroHow it worksFind a modelCompare modelsPricing

The best open LLMs you can run on 12 GB of VRAM

Open LLMs that fit in 12 GB of VRAM at their default quant — the sweet spot for an RTX 3060 12 GB, 4070, or 6700 XT. Ranked by popularity/recognition within the set that fits, with the honest $0-local, rent-a-GPU, and your-own-API-key cost for each. We guarantee the fit; you judge which one you like best.

How this is ranked: Objective fit filter only (fills the gap between the 8 and 16 GB tiers). 'Best' means 'runs on a 12 GB card.' VRAM is engine-computed; ordering is by popularity/recognition (a real signal), not a quality ranking we'd have to invent.

Showing the top 40 of 304. See all →

More: all "best" lists · Outcome Lab · all models

Open the free Spanvero advisor → · Honest, $0-markup. © 2026 Cynosure LLC.

The weekly price index

A short email of real AI price moves, straight from the daily log — no hype. We're collecting the list now; the first issue goes out when it opens. Unsubscribe with one click.