Models / GLM 5.2 / API pricing

GLM 5.2 API pricing

Live per-token API pricing across providers, synced from OpenRouter. Compare input, output, and cache-read rates, throughput, latency, and uptime in one place - then estimate what your workload actually costs.

Refreshed about 10 hours ago via OpenRouter - some pricing may be stale.

CHEAPEST PROVIDER

Relace - $0.17/M input, $12.00/M output , $0.150/M cache

Per 1M tokens, USD. Verify on the provider's site before committing.

See all providers below

Per-provider pricing

Live per-provider pricing, throughput and uptime - refreshed about 10 hours ago via OpenRouter. Click a column to sort.

some pricing may be stale - last verified 2026-10-08

Provider Type Input $/M Output $/M Cache $/M Tok/s Latency Uptime Value
Relace
API 0.17 12.00 0.150 - - 100.00% best uptime
DigitalOcean
API 0.70 2.20 0.105 - - 100.00%
DigitalOcean stale
API 0.70 2.20 0.105 - - -
CoreWeave
API 0.76 2.42 0.140 - - 100.00%
Nous Portal stale
API 0.95 2.99 - - - -
API 1.26 3.00 0.220 - - 100.00%
API 1.40 4.40 0.260 - - 100.00%
Parasail
API 1.40 4.40 0.260 - - 100.00%
Venice
API 1.40 4.40 0.260 - - 100.00%
Z.ai stale
API 1.40 4.40 - - - -
Mistral
API 1.54 4.84 0.154 - - 100.00%
BaseTen
API 2.10 6.60 0.210 - - 100.00%
Baidu
API 2.25 7.88 0.560 - - 100.00%
Alibaba
API 2.31 7.26 0.462 - - 100.00%
Decart
API 2.50 9.00 0.540 - - 100.00%
Morph
API 0.19 3.55 0.137 - - 99.98%
API 0.19 4.30 0.180 - - 99.98%
InferenceNet
API 0.18 4.40 0.120 - - 99.95%
Z.AI
API 1.40 4.40 0.260 - - 99.92%
AtlasCloud
API 0.94 2.95 0.174 - - 99.91%
SiliconFlow
API 1.19 3.74 0.221 - - 99.67%
Together
API 1.40 4.40 0.260 - - 99.45%
Cloudflare
API 1.18 4.40 0.260 - - 99.37%
API 1.40 4.40 0.260 - - 99.30%
Inceptron
API 1.39 4.39 0.250 - - 99.10%
DeepInfra
API 0.56 1.80 0.105 - - 98.61%
StreamLake risky
API 0.64 2.02 0.119 - - 94.77%
Novita risky
API 0.65 2.04 0.121 - - 94.34%
Nebius avoid
API 1.40 4.40 0.150 - - 33.36%
Sub - - - - - - $10.00/mo Coding Plan Lite
Sub - - - - - - $10.00/mo Go ($5 first month)
Sub - - - - - - $20.00/mo Pro
Sub - - - - - - $30.00/mo Coding Plan Pro
Sub - - - - - - $80.00/mo Coding Plan Max
Sub - - - - - - $100.00/mo Max

Default order: throughput among 95%+ uptime providers, then latency; subscriptions last. Sort by any column. Subscription rows show $/mo in the Value column - per-token columns are "-". Affiliate links are marked sponsored / nofollow. Confirm current pricing on the provider's site before committing.

PROGRAMMATIC

Get this data as JSON

Same provider table, machine-readable. No auth, no rate limit beyond the edge cache.

curl -s https://tokenstead.ai/models/glm-5-2/pricing.json

JSON: model metadata, cheapest_api, providers[], subscriptions[], verified_at. Unit is USD per 1M tokens.

NEXT STEP

Estimate what GLM 5.2 costs for your workload

Paste your prompt, set requests/day, and see monthly cost against self-hosting on your own GPU.

Open the calculator →
PRICE HISTORY

Inference cost over time

Data accumulates from the first daily sync - longer ranges populate over time. Prices come from OpenRouter snapshots, not a historical API.

Loading price history...