Gemini 3.7 Flash
workstationGoogle’s most intelligent workhorse model - the third Flash release of summer 2026 (3.5 Flash May, 3.6 Flash Jul, 3.7 Flash Aug 13). 1M context, 64K output, multimodal (text, image, video, audio, PDF) in, text out. Customizable thinking; tool use, search, and computer use.
Pricing: introductory $0.75/1M input, $3.75/1M output through Dec 31, 2026, then $1.50/$7.50. Batch 50% off.
Agents on Rails benchmark (Aug 2026, Le Mans round). 71.4% accuracy on 63 runs at $0.283 mean cost - the strongest accuracy-per-dollar below the top cluster, edging out Luna’s $0.014 cost-per-point tradeoff for mid-pack Rails work. API recall 27%.
- 1000k
- proprietary
- 🇺🇸 USA
- Aug 2026
Scores
Score per dollar
116 pts per $/M input
general_score (87) divided by cheapest input price ($0.75/M). Higher is better value. See live pricing.
Related models
Guides covering Gemini 3.7 Flash
Save your hardware and every model page answers the real question: will it run on your machine, and how fast?
Join free - save your rig →Or run it in the cloud
Live per-provider pricing, throughput and uptime - refreshed about 2 months ago via OpenRouter. Click a column to sort.
some pricing may be stale - last verified 2026-08-24
| Provider | Type | Input $/M | Output $/M | Cache $/M | Tok/s | Latency | Uptime | Value |
|---|---|---|---|---|---|---|---|---|
|
Google Gemini
stale
|
API | 0.75 | 3.75 | - | - | - | - | cheapest |
Default order: throughput among 95%+ uptime providers, then latency; subscriptions last. Sort by any column. Subscription rows show $/mo in the Value column - per-token columns are "-". Affiliate links are marked sponsored / nofollow. Confirm current pricing on the provider's site before committing.
Detailed API pricing page + JSON endpoint →
See who runs Google in production →
Inference cost over time
Data accumulates from the first daily sync - longer ranges populate over time. Prices come from OpenRouter snapshots, not a historical API.