Browse AI models
18 models tracked - grouped by family, best families first
Jeff
mstrasser (firelex) 1 model best 82Nemotron
๐บ๐ธ NVIDIA 2 models best 80Ornith
๐บ๐ธ Ornith AI 1 model best 74Qwen
๐จ๐ณ Alibaba 2 models best 65Gliner
๐บ๐ธ Fastino 1 model best 64Llama
๐บ๐ธ Meta 3 models best 62
Llama 3.2 1B Instruct
Sep 25, 2024
no marks
1.2B
128k ctx
1GB min RAM
llama 3
2 var
$0.03/M in
35
1296 pts/$
Llama 3.2 3B Instruct
Sep 25, 2024
no marks
3.2B
128k ctx
2GB min RAM
llama 3
2 var
$0.05/M in
50
1000 pts/$
Llama 3.1 8B Instruct
Jul 23, 2024
no marks
8.0B
128k ctx
5GB min RAM
llama 3
3 var
$0.02/M in
62
3100 pts/$
Other labs (Kimi ยท LFM ยท Command R+)
๐บ๐ธ Liquid AI 1 model best 58Gemma
๐บ๐ธ Google 4 models best 55
Gemma 3 270M
Aug 14, 2025
no marks
270M
32k ctx
1GB min RAM
other
2 var
25
Gemma 3n E2B
Jun 26, 2025
no marks
2.0B
32k ctx
6GB min RAM
other
2 var
48
Gemma 3 1B
Feb 27, 2025
no marks
1.0B
32k ctx
1GB min RAM
other
2 var
38
Gemma 3 4B
Feb 27, 2025
no marks
4.3B
128k ctx
3GB min RAM
other
2 var
$0.05/M in
55
1100 pts/$
Phi
๐บ๐ธ Microsoft 1 model best 52Laya
ConvAI Innovations 1 model best 0Openjev
Undisclosed 1 model best 0FAQ
What is the cheapest LLM API?
Open-weight models like DeepSeek, Qwen, and GLM are the cheapest LLM APIs - often under $1 per million input tokens. Sort the table by best value to rank quality per dollar.
Which LLM has the best coding score?
Coding scores vary by benchmark; filter the table by use case coding to see coding_score per model, or open any model's page for its benchmark breakdown.
How do I compare LLM pricing?
This table shows per-million-token input and output prices across providers. Open a model's pricing page for live rates, cached-input discounts, and price history.