Qwen3.8-Omni-Flash API pricing
Live per-token API pricing across providers, synced from OpenRouter. Compare input, output, and cache-read rates, throughput, latency, and uptime in one place - then estimate what your workload actually costs.
Refreshed 4 days ago via OpenRouter.
Alibaba Cloud Model Studio - $0.15/M input, $0.47/M output , $0.016/M cache
Per 1M tokens, USD. Verify on the provider's site before committing.
Per-provider pricing
Live per-provider pricing, throughput and uptime - refreshed 4 days ago via OpenRouter. Click a column to sort.
| Provider | Type | Input $/M | Output $/M | Cache $/M | Tok/s | Latency | Uptime | Value |
|---|---|---|---|---|---|---|---|---|
| API | 0.15 | 0.47 | 0.016 | - | - | - | cheapest |
Default order: throughput among 95%+ uptime providers, then latency; subscriptions last. Sort by any column. Subscription rows show $/mo in the Value column - per-token columns are "-". Affiliate links are marked sponsored / nofollow. Confirm current pricing on the provider's site before committing.
Get this data as JSON
Same provider table, machine-readable. No auth, no rate limit beyond the edge cache.
curl -s https://tokenstead.ai/models/qwen3-8-omni-flash/pricing.json
JSON: model metadata, cheapest_api, providers[], subscriptions[], verified_at. Unit is USD per 1M tokens.
Estimate what Qwen3.8-Omni-Flash costs for your workload
Paste your prompt, set requests/day, and see monthly cost against self-hosting on your own GPU.
Inference cost over time
Data accumulates from the first daily sync - longer ranges populate over time. Prices come from OpenRouter snapshots, not a historical API.