Can Used NVIDIA RTX 3090 24GB run Llama, Qwen & DeepSeek? 352 models that fit
352 of the 482 models in the Spanvero catalog fit Used NVIDIA RTX 3090 24GB's 24 GB VRAM (at a sensible quant, 16k context). For each: run it locally ($0 compute + electricity), rent an equivalent GPU ($0 markup, as of 2026-07-09), or pay per-token via your own API key (as of 2026-08-10).
Three honest ways to run each model on Used NVIDIA RTX 3090 24GB
Run it locally: $0 in compute — you pay only electricity (~350 W under load on this card). Local is real money, never a fake "$0".
Rent an equivalent GPU: from a $0-markup vendor rate (as of 2026-07-09) — you rent on your own account and pay the vendor directly; we never resell compute.
Skip the box: run the same model through your own API key, paying per million tokens (prices as of 2026-08-10).
What fits Used NVIDIA RTX 3090 24GB (24 GB VRAM)
352 of the 482 notable models in the Spanvero catalog fit Used NVIDIA RTX 3090 24GB at a sensible quant (context capped at 16k for the estimate). Most capable first:
Laguna XS.2 (poolside, 33.4B) — needs ~24 GB at Q4_K_M: run it locally for $0 compute + ~$0.6045/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.15/1M via your own API key (last-known).
Laguna XS 2.1 NVFP4 (poolside, 33.4B) — needs ~24 GB at Q4_K_M: run it locally for $0 compute + ~$0.6045/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.37/1M via your own API key (size estimate).
Laguna XS 2.1 (poolside, 33.4B) — needs ~24 GB at Q4_K_M: run it locally for $0 compute + ~$0.6045/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.09/1M via your own API key.
sarvam 30b (sarvamai, 32.2B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5871/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.36/1M via your own API key (size estimate).
llm jp 4 32b a3b thinking (llm-jp, 32.1B) — needs ~23 GB at Q4_K_M: run it locally for $0 compute + ~$0.5856/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.36/1M via your own API key (size estimate).
NVIDIA Nemotron 3 Nano 30B A3B BF16 (nvidia, 31.6B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5783/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.35/1M via your own API key (size estimate).
Nemotron Cascade 2 30B A3B (nvidia, 31.6B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5783/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.35/1M via your own API key (size estimate).
Qwen3 30B A3B (Qwen, 30.5B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5622/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.31/1M via your own API key.
Qwen3 Coder 30B A3B Instruct (Qwen, 30.5B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5622/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.17/1M via your own API key.
Qwen3 30B A3B Instruct 2507 (Qwen, 30.5B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5622/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.12/1M via your own API key.
Qwen3 30B A3B Thinking 2507 (Qwen, 30.5B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5622/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$1.30/1M via your own API key.
Tongyi DeepResearch 30B A3B (Alibaba-NLP, 30.5B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5622/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.34/1M via your own API key (size estimate).
Qwen3 30B A3B Base (Qwen, 30.5B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5622/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.34/1M via your own API key (size estimate).
lynx instruct 30b (bineric, 30.5B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5622/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.34/1M via your own API key (size estimate).
North Mini Code 1.0 (CohereLabs, 30.5B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5622/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.34/1M via your own API key (size estimate).
granite 4.1 30b (ibm-granite, 28.9B) — needs ~24 GB at Q4_K_M: run it locally for $0 compute + ~$0.5384/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.33/1M via your own API key (size estimate).
Gemma 2 27B Instruct (Google, 27B) — needs ~22 GB at Q4_K_M: run it locally for $0 compute + ~$0.5099/1M in electricity, rent 3× NVIDIA RTX 3060 12GB from $0.18/hr ($0 markup), or ~$0.65/1M via your own API key.
Trinity Mini (arcee-ai, 26.1B) — needs ~19 GB at Q4_K_M: run it locally for $0 compute + ~$0.4963/1M in electricity, rent 2× NVIDIA RTX 3060 12GB from $0.12/hr ($0 markup), or ~$0.10/1M via your own API key (last-known).
LFM2 24B A2B (LiquidAI, 23.8B) — needs ~18 GB at Q4_K_M: run it locally for $0 compute + ~$0.461/1M in electricity, rent 2× NVIDIA RTX 3060 12GB from $0.12/hr ($0 markup), or ~$0.29/1M via your own API key (size estimate).
Mistral Small 3 (24B, 2501) (Mistral AI, 23.6B) — needs ~20 GB at Q4_K_M: run it locally for $0 compute + ~$0.4579/1M in electricity, rent 2× NVIDIA RTX 3060 12GB from $0.12/hr ($0 markup), or ~$0.07/1M via your own API key.
The honest cost of owning Used NVIDIA RTX 3090 24GB
The full ownership math for this card — its dated street price, 3-year amortization, electricity, and the own-vs-rent break-evens — lives on the dedicated NVIDIA RTX 3090 24GB GPU page (linked under "Keep exploring" below), so those numbers are published once, in one place.
Too big for Used NVIDIA RTX 3090 24GB — rent or use an API instead
These need more than the 24 GB VRAM on this card. Closest first — you can still run them on a rented GPU ($0 markup) or via your own API key:
Yi-1.5-34B-Chat (01.AI, 34.4B) — needs ~25 GB; rent 3× NVIDIA RTX 3060 12GB from $0.18/hr, or ~$0.38/1M via your own API key (size estimate).
Qwen3-32B (Alibaba, 32.8B) — needs ~25 GB; rent 3× NVIDIA RTX 3060 12GB from $0.18/hr, or ~$0.18/1M via your own API key.
granite 4.0 h small (ibm-granite, 32.2B) — needs ~25 GB; rent 3× NVIDIA RTX 3060 12GB from $0.18/hr, or ~$0.36/1M via your own API key (size estimate).
Qwen2.5-Coder 32B Instruct (Alibaba, 32B) — needs ~25 GB; rent 3× NVIDIA RTX 3060 12GB from $0.18/hr, or ~$0.83/1M via your own API key.
Gemma 4 31B IT NVFP4 (nvidia, 20.9B) — needs ~25 GB; rent 3× NVIDIA RTX 3060 12GB from $0.18/hr, or ~$0.27/1M via your own API key (size estimate).
Olmo 3 1125 32B (allenai, 32.2B) — needs ~26 GB; rent 3× NVIDIA RTX 3060 12GB from $0.18/hr, or ~$0.36/1M via your own API key (size estimate).
A short email of real AI price moves, straight from the daily log — no hype. We're collecting the list now; the first issue goes out when it opens. Unsubscribe with one click.