price of compute

The price of a token

API list prices for frontier models, in USD per million tokens — standard tier, text, base context. Each number is verified against the lab’s official pricing page (date shown); the daily series is in the free API. More labs get added as they’re verified — never before.

Input vs output $/MTok · log-log · 5×/6× guides
$3$10$0.3$1$3input →output ↑5×6×gemini-3.1-pro-previewgemini-3.5-flashgemini-3.6-flashgemini-3.5-flash-lite

The dashed rays are constant output÷input ratios — nearly the whole market prices output at 5–6× input. Distance along a ray is price class; distance off it is pricing strategy.

The ladder · every model, one scale
$0.3$1$3$10gemini-3.5-flash-lite: $0.30 in · $2.50 out per MTokgemini-3.5-flash-litegemini-3.6-flash: $1.50 in · $7.50 out per MTokgemini-3.6-flashgemini-3.5-flash: $1.50 in · $9.00 out per MTokgemini-3.5-flashgemini-3.1-pro-preview: $2.00 in · $12.00 out per MTokgemini-3.1-pro-previewinputoutput · $/MTok, log scale

Google

ModelInput $/MTokOutput $/MTokOut÷InBlended 3:1Verified
gemini-3.1-pro-preview$2.00$12.006.0×$4.502026-08-10
gemini-3.5-flash$1.50$9.006.0×$3.382026-08-10
gemini-3.6-flash$1.50$7.505.0×$3.002026-08-10
gemini-3.5-flash-lite$0.30$2.508.3×$0.852026-08-10

Last snapshot: 2026-08-11 06:05 UTC · List prices, not negotiated rates; cached-input, batch, and long-context tiers vary by lab and are noted where they differ. Same rule as the GPU side: a missing number beats a wrong one.