SpanveroCheck my hardwareCompare GPUsModels by VRAMCompare costsHow it works

All AI models — compare three cost routes

Browse 604 open models by size, publisher, hardware budget, or use case. Each page separates calculated local requirements, provider-published rented-GPU rates, and dated API rates from estimates.

By size: Flagship (80B+) · Large (34–80B) · Medium (14–34B) · Small (4–14B) · Tiny (under 4B)

By your GPU: 8 GB of VRAM · 16 GB of VRAM · 24 GB of VRAM · 48 GB of VRAM

By publisher: Qwen (Alibaba) · Meta Llama · DeepSeek · Google (Gemma) · Microsoft (Phi) · Mistral AI · NVIDIA · OpenAI (gpt-oss) · IBM Granite · Z.ai (GLM) · Moonshot AI (Kimi) · MiniMax · Cohere · Ai2 (OLMo) · Nous Research · Liquid AI · TII (Falcon) · EleutherAI · InternLM · Xiaomi (MiMo)

By use case: Best by use · 24 GB VRAM · 12 GB VRAM · Laptop LLMs · Speech to text · Compare · Cost calculator

Flagship (80B+)

The biggest open models — multi-GPU or API territory. see all 93 →

Large (34–80B)

Top quality that still fits a single big rented GPU. see all 52 →

Medium (14–34B)

The single-GPU sweet spot — strong and self-hostable. see all 114 →

Small (4–14B)

Runs on a good laptop or consumer GPU. see all 189 →

Tiny (under 4B)

Edge / on-device — runs almost anywhere. see all 156 →

Compare models head-to-head → · LLM cost calculator →

Open the free LLM cost calculator → · Rate table dated 2026-08-24. $0 markup, your own accounts, we never resell compute. © 2026 Cynosure LLC.

The weekly price index

A short email of real AI price moves, straight from the daily log — no hype. We're collecting the list now; the first issue goes out when it opens. Unsubscribe with one click.