Models / GPT-5.6 Terra / Calculator

GPT-5.6 Terra cost & VRAM calculator

What does GPT-5.6 Terra cost to run for your workload, and can you run it on your own hardware? Set your workload below - we compute per-provider API cost live and tell you honestly whether local hardware can run it.

LOCAL ISN'T PRACTICAL

This is where owning stops making sense

GPT-5.6 Terra is a large-billion-parameter model. No rig an individual can buy runs it - so unlike a workstation GPU model, there's no break-even to compute. The honest answer for nearly everyone is the per-provider API cost below.

Efficient proprietary dense model from OpenAI, the mid-tier sibling of GPT-5.6 Sol. Same 1,050,000-token context and 128K max output as Sol. Text + image in, text out. Reasoning effort spans none through max; medium is default. Pricing: $2/1M input, $12/1M output. Agents on Rails benchmark (Aug 2026, Le Mans round). 77.8% accuracy on 63 runs - 49 of 63, at $0.197 mean cost and a 3m 02s median, the fastest in the entire 16-model field. OpenAI’s three models (Sol 84.1%, Terra 77.8%, Luna 73%) score in exactly price order, all in the fastest third. Terra is the “20 cents a run” sweet spot: near-frontier accuracy at a fraction of the flagship’s cost.

Your workload

GPT-5.6 Terra runs an always-on thinking mode. Reasoning (thinking) tokens are billed at the output rate ($15.00/M), so count them here to see the thinking portion of your bill.

API cost for your workload

Provider Rate ($/1M) Monthly cost
OpenAI may be stale $2.00 in · $12.00 out $0.05 cheapest
DigitalOcean may be stale $2.00 in · $12.00 out · $0.20 cache $0.05 cheapest

Monthly cost is an estimate from list prices and your workload - verify against the provider before committing. Cached fraction applies the cache rate to that share of input.

Can you run it locally?

NO - NO INDIVIDUAL RIG RUNS IT

GPT-5.6 Terra has no published quantization that fits a rig one person can buy, so there's no local-hardware recommendation and no break-even to compute. The honest answer is the per-provider API cost above.

Efficient proprietary dense model from OpenAI, the mid-tier sibling of GPT-5.6 Sol. Same 1,050,000-token context and 128K max output as Sol. Text + image in, text out. Reasoning effort spans none through max; medium is default. Pricing: $2/1M input, $12/1M output. Agents on Rails benchmark (Aug 2026, Le Mans round). 77.8% accuracy on 63 runs - 49 of 63, at $0.197 mean cost and a 3m 02s median, the fastest in the entire 16-model field. OpenAI’s three models (Sol 84.1%, Terra 77.8%, Luna 73%) score in exactly price order, all in the fastest third. Terra is the “20 cents a run” sweet spot: near-frontier accuracy at a fraction of the flagship’s cost.

See the model card for the full architecture notes and any cloud subscription plans.

Full model card API pricing table Generic token calculator