Model
Fit
Speed
Memory
Scores / Run
Runs comfortably
6.2GB
14k cap
Runs comfortably
11
tok/s
ⓘ estimated
6
/ 16GB
88
/
89
/
85
Runs comfortably
11 tok/s
ⓘ estimated
6/16GB
9.8GB
9k cap
Runs comfortably
8
tok/s
ⓘ estimated
10
/ 16GB
88
/
89
/
85
Runs comfortably
8 tok/s
ⓘ estimated
10/16GB
Runs comfortably
8 tok/s
ⓘ estimated
10/16GB
Runs comfortably
10 tok/s
ⓘ estimated
8/16GB
Runs comfortably
11 tok/s
ⓘ estimated
7/16GB
Runs comfortably
12 tok/s
ⓘ estimated
6/16GB
Runs comfortably
9 tok/s
ⓘ estimated
9/16GB
Runs comfortably
8 tok/s
ⓘ estimated
9/16GB
Runs comfortably
10 tok/s
ⓘ estimated
7/16GB
Runs comfortably
6 tok/s
ⓘ estimated
13/16GB
Runs comfortably
13 tok/s
ⓘ estimated
5/16GB
Runs comfortably
9 tok/s
ⓘ estimated
9/16GB
Runs comfortably
10 tok/s
ⓘ estimated
8/16GB
Runs comfortably
6 tok/s
ⓘ estimated
14/16GB
Runs comfortably
6 tok/s
ⓘ estimated
13/16GB
Runs comfortably
10 tok/s
ⓘ estimated
7/16GB
Runs comfortably
24 tok/s
ⓘ estimated
3/16GB
Runs comfortably
19 tok/s
ⓘ estimated
4/16GB
Runs comfortably
28 tok/s
ⓘ estimated
3/16GB
Runs comfortably
9 tok/s
ⓘ estimated
8/16GB
Llama 3.1 8B Instruct
Q4_K_M
4.5GB
23k cap
Runs comfortably
14
tok/s
ⓘ estimated
5
/ 16GB
52
/
55
/
62
Runs comfortably
14 tok/s
ⓘ estimated
5/16GB
Runs comfortably
15 tok/s
ⓘ estimated
5/16GB
Runs comfortably
22 tok/s
ⓘ estimated
3/16GB
3.5GB
52k cap
Runs comfortably
20
tok/s
ⓘ estimated
4
/ 16GB
35
/
38
/
50
Runs comfortably
20 tok/s
ⓘ estimated
4/16GB
Llama 3.2 3B Instruct
Q4_K_M
2.0GB
58k cap
Runs comfortably
31
tok/s
ⓘ estimated
2
/ 16GB
35
/
38
/
50
Runs comfortably
31 tok/s
ⓘ estimated
2/16GB
Runs tight
Runs tight
18 tok/s
ⓘ estimated
14/16GB
Gemma 4 26B A4B
Q4_K_M
13.0GB
ctx headroom ~0
Runs tight
19
tok/s
ⓘ estimated
13
/ 16GB
85
/
88
/
84
Runs tight
19 tok/s
ⓘ estimated
13/16GB
Runs tight
24 tok/s
ⓘ estimated
16/16GB
Runs tight
6 tok/s
ⓘ estimated
15/16GB
Runs tight
5 tok/s
ⓘ estimated
15/16GB
16.0GB
ctx headroom ~0
Runs tight
5
tok/s
ⓘ estimated
16
/ 16GB
52
/
55
/
62
Runs tight
5 tok/s
ⓘ estimated
16/16GB
Coding alternatives
Hosted coding assistants for comparison - per seat
Cursor
Hobby
AI-first code editor. Uses frontier models (GPT-4, Claude). $20/user Business...
$0.0
GitHub Copilot
Individual
AI coding assistant by GitHub. GPT-4o and Claude-based completions. $19/seat ...
$10.0
GitHub Copilot
Business
AI coding assistant by GitHub. GPT-4o and Claude-based completions. $19/seat ...
$19.0
Cursor
Pro
AI-first code editor. Uses frontier models (GPT-4, Claude). $20/user Business...
$20.0
GitHub Copilot
Enterprise
AI coding assistant by GitHub. GPT-4o and Claude-based completions. $19/seat ...
$39.0
Cursor
Business
AI-first code editor. Uses frontier models (GPT-4, Claude). $20/user Business...
$40.0
Per-token monthly figures are estimates at typical usage. Prices verified periodically - always confirm current pricing on the provider's website.
Sign in with GitHub to save this rig and get weekly model updates
Sign in with GitHub