Claude Opus 4.8
premierProprietary dense model from Anthropic, the previous flagship before Opus 5. Parameter count is undisclosed. Native 1M-token context. Adaptive thinking with effort levels; strong agentic coding and long-horizon capability.
- Pricing: $5/1M input, $25/1M output.
Agents on Rails benchmark (Aug 2026, Le Mans round). 79.4% accuracy on 63 runs - tied with GLM 5.3 and the fastest model at its score tier (3m 36s median, ~15% of Opus 5βs time). API recall 15.9%, the third-lowest in the field. Superseded by Opus 5 (92.1%) on this board, but the accuracy-per-minute is the best of any model above 79%.
- 1000k
- proprietary
- πΊπΈ USA
- May 2026
Scores
Score per dollar
19 pts per $/M input
general_score (93) divided by cheapest input price ($5.00/M). Higher is better value. See live pricing.
Related models
Guides covering Claude Opus 4.8
Save your hardware and every model page answers the real question: will it run on your machine, and how fast?
Join free - save your rig βOr run it in the cloud
Live per-provider pricing, throughput and uptime - refreshed about 2 months ago via OpenRouter. Click a column to sort.
some pricing may be stale - last verified 2026-08-12
| Provider | Type | Input $/M | Output $/M | Cache $/M | Tok/s | Latency | Uptime | Value |
|---|---|---|---|---|---|---|---|---|
|
Anthropic
stale
|
API | 5.00 | 25.00 | - | - | - | - | cheapest |
Default order: throughput among 95%+ uptime providers, then latency; subscriptions last. Sort by any column. Subscription rows show $/mo in the Value column - per-token columns are "-". Affiliate links are marked sponsored / nofollow. Confirm current pricing on the provider's site before committing.
Detailed API pricing page + JSON endpoint β
Inference cost over time
Data accumulates from the first daily sync - longer ranges populate over time. Prices come from OpenRouter snapshots, not a historical API.