Browse AI models
39 models tracked - grouped by family, best families first
Qwen Image
π¨π³ Alibaba 1 model best 90Pokee
πΊπΈ Pokee AI 1 model best 88Qwen
π¨π³ Alibaba 8 models best 87
Qwen3.8-27B
Aug 5, 2026
27.0B
262k ctx
8GB min RAM
apache 2.0
4 var
85
Qwen3.6 35B A3B
MoE
Apr 22, 2026
text + metadata
35.0B
3.0B active
262k ctx
20GB min RAM
apache 2.0
2 var
$0.05/M in
87
1740 pts/$
Qwen3.6 27B
Nov 1, 2025
text + metadata
27.0B
128k ctx
17GB min RAM
apache 2.0
2 var
$0.30/M in
82
273 pts/$
Qwen3 14B
Apr 28, 2025
text + metadata
14.7B
128k ctx
9GB min RAM
apache 2.0
2 var
$0.10/M in
70
700 pts/$
Qwen3 30B A3B
MoE
Apr 28, 2025
text + metadata
30.5B
3.0B active
128k ctx
16GB min RAM
apache 2.0
2 var
$0.12/M in
72
600 pts/$
Qwen3 32B
Apr 28, 2025
text + metadata
32.8B
128k ctx
20GB min RAM
apache 2.0
4 var
$0.08/M in
78
975 pts/$
Qwen3 4B
Apr 28, 2025
text + metadata
4.0B
32k ctx
3GB min RAM
apache 2.0
1 var
56
Qwen3 8B
Apr 28, 2025
text + metadata
8.2B
128k ctx
5GB min RAM
apache 2.0
2 var
$0.12/M in
65
556 pts/$
Ornith
πΊπΈ Ornith AI 2 models best 86Gemma
πΊπΈ Google 10 models best 85
Gemma 4 26B A4B
MoE
Apr 30, 2026
no marks
25.2B
3.8B active
256k ctx
15GB min RAM
other
4 var
$0.04/M in
84
2000 pts/$
Gemma 4 31B
Apr 2, 2026
no marks
30.7B
256k ctx
19GB min RAM
apache 2.0
5 var
$0.08/M in
85
1063 pts/$
Gemma 3 270M
Aug 14, 2025
no marks
270M
32k ctx
1GB min RAM
other
2 var
25
Gemma 3n E2B
Jun 26, 2025
no marks
2.0B
32k ctx
6GB min RAM
other
2 var
48
Gemma 3n E4B
Jun 26, 2025
no marks
4.0B
32k ctx
8GB min RAM
other
2 var
58
Gemma 4 12B
Jun 1, 2025
no marks
12.0B
128k ctx
7GB min RAM
other
2 var
72
Gemma 3 12B
Feb 27, 2025
no marks
12.2B
128k ctx
8GB min RAM
other
2 var
$0.05/M in
68
1360 pts/$
Gemma 3 1B
Feb 27, 2025
no marks
1.0B
32k ctx
1GB min RAM
other
2 var
38
Gemma 3 27B
Feb 27, 2025
no marks
27.2B
128k ctx
17GB min RAM
other
2 var
$0.08/M in
72
900 pts/$
Gemma 3 4B
Feb 27, 2025
no marks
4.3B
128k ctx
3GB min RAM
other
2 var
$0.05/M in
55
1100 pts/$
Jeff
mstrasser (firelex) 1 model best 82Llama
πΊπΈ Meta 4 models best 80
Llama 3.3 70B Instruct
Dec 6, 2024
no marks
70.6B
128k ctx
22GB min RAM
llama 3
4 var
$0.10/M in
80
800 pts/$
Llama 3.2 1B Instruct
Sep 25, 2024
no marks
1.2B
128k ctx
1GB min RAM
llama 3
2 var
$0.03/M in
35
1296 pts/$
Llama 3.2 3B Instruct
Sep 25, 2024
no marks
3.2B
128k ctx
2GB min RAM
llama 3
2 var
$0.05/M in
50
1000 pts/$
Llama 3.1 8B Instruct
Jul 23, 2024
no marks
8.0B
128k ctx
5GB min RAM
llama 3
3 var
$0.02/M in
62
3100 pts/$
Nemotron
πΊπΈ NVIDIA 2 models best 80Flux
π©πͺ Black Forest Labs 1 model best 78Phi
πΊπΈ Microsoft 2 models best 70Mistral
π«π· Mistral AI 2 models best 68Gliner
πΊπΈ Fastino 1 model best 64Other labs (Kimi Β· LFM Β· Command R+)
πΊπΈ Liquid AI 1 model best 58Laya
ConvAI Innovations 1 model best 0Openjev
Undisclosed 1 model best 0FAQ
What is the cheapest LLM API?
Open-weight models like DeepSeek, Qwen, and GLM are the cheapest LLM APIs - often under $1 per million input tokens. Sort the table by best value to rank quality per dollar.
Which LLM has the best coding score?
Coding scores vary by benchmark; filter the table by use case coding to see coding_score per model, or open any model's page for its benchmark breakdown.
How do I compare LLM pricing?
This table shows per-million-token input and output prices across providers. Open a model's pricing page for live rates, cached-input discounts, and price history.