GLM 5.2 API pricing
Live per-token API pricing across providers, synced from OpenRouter. Compare input, output, and cache-read rates, throughput, latency, and uptime in one place - then estimate what your workload actually costs.
Refreshed about 10 hours ago via OpenRouter - some pricing may be stale.
Relace - $0.17/M input, $12.00/M output , $0.150/M cache
Per 1M tokens, USD. Verify on the provider's site before committing.
Per-provider pricing
Live per-provider pricing, throughput and uptime - refreshed about 10 hours ago via OpenRouter. Click a column to sort.
some pricing may be stale - last verified 2026-10-08
| Provider | Type | Input $/M | Output $/M | Cache $/M | Tok/s | Latency | Uptime | Value |
|---|---|---|---|---|---|---|---|---|
|
Relace
|
API | 0.17 | 12.00 | 0.150 | - | - | 100.00% | best uptime |
|
DigitalOcean
|
API | 0.70 | 2.20 | 0.105 | - | - | 100.00% | |
|
DigitalOcean
stale
|
API | 0.70 | 2.20 | 0.105 | - | - | - | |
|
CoreWeave
|
API | 0.76 | 2.42 | 0.140 | - | - | 100.00% | |
|
Nous Portal
stale
|
API | 0.95 | 2.99 | - | - | - | - | |
| API | 1.26 | 3.00 | 0.220 | - | - | 100.00% | ||
| API | 1.40 | 4.40 | 0.260 | - | - | 100.00% | ||
|
Parasail
|
API | 1.40 | 4.40 | 0.260 | - | - | 100.00% | |
|
Venice
|
API | 1.40 | 4.40 | 0.260 | - | - | 100.00% | |
|
Z.ai
stale
|
API | 1.40 | 4.40 | - | - | - | - | |
|
Mistral
|
API | 1.54 | 4.84 | 0.154 | - | - | 100.00% | |
|
BaseTen
|
API | 2.10 | 6.60 | 0.210 | - | - | 100.00% | |
|
Baidu
|
API | 2.25 | 7.88 | 0.560 | - | - | 100.00% | |
|
Alibaba
|
API | 2.31 | 7.26 | 0.462 | - | - | 100.00% | |
|
Decart
|
API | 2.50 | 9.00 | 0.540 | - | - | 100.00% | |
|
Morph
|
API | 0.19 | 3.55 | 0.137 | - | - | 99.98% | |
| API | 0.19 | 4.30 | 0.180 | - | - | 99.98% | ||
|
InferenceNet
|
API | 0.18 | 4.40 | 0.120 | - | - | 99.95% | |
|
Z.AI
|
API | 1.40 | 4.40 | 0.260 | - | - | 99.92% | |
|
AtlasCloud
|
API | 0.94 | 2.95 | 0.174 | - | - | 99.91% | |
|
SiliconFlow
|
API | 1.19 | 3.74 | 0.221 | - | - | 99.67% | |
|
Together
|
API | 1.40 | 4.40 | 0.260 | - | - | 99.45% | |
|
Cloudflare
|
API | 1.18 | 4.40 | 0.260 | - | - | 99.37% | |
| API | 1.40 | 4.40 | 0.260 | - | - | 99.30% | ||
|
Inceptron
|
API | 1.39 | 4.39 | 0.250 | - | - | 99.10% | |
|
DeepInfra
|
API | 0.56 | 1.80 | 0.105 | - | - | 98.61% | |
|
StreamLake
risky
|
API | 0.64 | 2.02 | 0.119 | - | - | 94.77% | |
|
Novita
risky
|
API | 0.65 | 2.04 | 0.121 | - | - | 94.34% | |
|
Nebius
avoid
|
API | 1.40 | 4.40 | 0.150 | - | - | 33.36% | |
| Sub | - | - | - | - | - | - | $10.00/mo Coding Plan Lite | |
| Sub | - | - | - | - | - | - | $10.00/mo Go ($5 first month) | |
| Sub | - | - | - | - | - | - | $20.00/mo Pro | |
| Sub | - | - | - | - | - | - | $30.00/mo Coding Plan Pro | |
| Sub | - | - | - | - | - | - | $80.00/mo Coding Plan Max | |
| Sub | - | - | - | - | - | - | $100.00/mo Max |
Default order: throughput among 95%+ uptime providers, then latency; subscriptions last. Sort by any column. Subscription rows show $/mo in the Value column - per-token columns are "-". Affiliate links are marked sponsored / nofollow. Confirm current pricing on the provider's site before committing.
Get this data as JSON
Same provider table, machine-readable. No auth, no rate limit beyond the edge cache.
curl -s https://tokenstead.ai/models/glm-5-2/pricing.json
JSON: model metadata, cheapest_api, providers[], subscriptions[], verified_at. Unit is USD per 1M tokens.
Estimate what GLM 5.2 costs for your workload
Paste your prompt, set requests/day, and see monthly cost against self-hosting on your own GPU.
Inference cost over time
Data accumulates from the first daily sync - longer ranges populate over time. Prices come from OpenRouter snapshots, not a historical API.