SpanveroCheck my hardwareCompare GPUsModels by VRAMCompare costsHow it works

The best open LLMs you can run on 24 GB of VRAM

Open LLMs calculated to need no more than 24 GB of VRAM at the stated default quant and context — the RTX 3090, 4090, and 7900 XTX tier. Ranked by popularity within that calculated fit set, with local and rented-GPU cost estimates. Verify the selected build and context on your exact machine.

How this is ranked: Transparent calculated fit filter. 'Best' means 'estimated to run within a 24 GB VRAM budget at the stated quant and context.' Ordering is by popularity/recognition, never a quality verdict; actual usage depends on runner overhead and settings.

Showing the top 40 of 426. See all →

Want every model that fits? All 426 models that run on 24 GB of VRAM →

More: all "best" lists · Outcome Lab · all models

Open the free LLM cost calculator → · $0 markup. © 2026 Cynosure LLC.

The weekly price index

A short email of real AI price moves, straight from the daily log — no hype. We're collecting the list now; the first issue goes out when it opens. Unsubscribe with one click.