Browse AI models
16 models tracked - grouped by family, best families first
GLM / Zhipu
π¨π³ Zhipu AI 2 models best 92Claude Sonnet
πΊπΈ Anthropic 2 models best 91DeepSeek
π¨π³ DeepSeek 5 models best 90
DeepSeek V4.1 Flash
MoE
Sep 10, 2026
552.0B
16.0B active
1000k ctx
265GB min RAM
mit
2 var
$0.02/M in
90
4500 pts/$
DeepSeek V4 Flash 0731
MoE
Jul 31, 2026
text + metadata
284.0B
13.0B active
1000k ctx
168GB min RAM
mit
2 var
$0.01/M in
90
7500 pts/$
DeepSeek V3.1 Terminus
MoE
Sep 29, 2025
text + metadata
685.0B
37.0B active
128k ctx
no local build
mit
$0.27/M in
82
304 pts/$
DeepSeek V3.2 Exp
MoE
Sep 29, 2025
text + metadata
685.0B
37.0B active
128k ctx
no local build
mit
$0.13/M in
83
618 pts/$
DeepSeek V3 0324
MoE
Mar 24, 2025
text + metadata
671.0B
37.0B active
128k ctx
188GB min RAM
mit
2 var
$0.24/M in
87
363 pts/$
GPT-5.6
πΊπΈ OpenAI 1 model best 90Ornith
πΊπΈ Ornith AI 1 model best 89Gemini
πΊπΈ Google 1 model best 87Qwen
π¨π³ Alibaba 1 model best 85Union Alpha (stealth, operator unconfirmed)
1 model best 85Mimo
π¨π³ Xiaomi 1 model best 76Nemotron
πΊπΈ NVIDIA 1 model best 74FAQ
What is the cheapest LLM API?
Open-weight models like DeepSeek, Qwen, and GLM are the cheapest LLM APIs - often under $1 per million input tokens. Sort the table by best value to rank quality per dollar.
Which LLM has the best coding score?
Coding scores vary by benchmark; filter the table by use case coding to see coding_score per model, or open any model's page for its benchmark breakdown.
How do I compare LLM pricing?
This table shows per-million-token input and output prices across providers. Open a model's pricing page for live rates, cached-input discounts, and price history.