SpanveroHow it worksFind a modelCompare modelsPricing

The best open LLMs you can run on 24 GB of VRAM

Open LLMs that fit in 24 GB of VRAM at their default quant — the RTX 3090 / 4090 / 7900 XTX tier where serious local models like 32B-class checkpoints become runnable. Ranked by popularity/recognition within the set that fits, with honest $0-local and rent-a-GPU costs. We guarantee the fit; you judge the quality.

How this is ranked: Objective fit filter only. 'Best' = 'runs on a 24 GB card.' VRAM is engine-computed; ordering by popularity/recognition within the fit set, never a quality verdict; the 24 GB VRAM hub lists every fitting model, largest first.

Showing the top 40 of 392. See all →

Want every model that fits? All 392 models that run on 24 GB of VRAM →

More: all "best" lists · Outcome Lab · all models

Open the free Spanvero advisor → · Honest, $0-markup. © 2026 Cynosure LLC.

The weekly price index

A short email of real AI price moves, straight from the daily log — no hype. We're collecting the list now; the first issue goes out when it opens. Unsubscribe with one click.