price of compute
Answers · data as of Aug 11 2026

Which LLM API is cheapest per token?

Among tracked frontier models, deepseek-v4-flash (deepseek) is currently the cheapest at a blended ~$0.18 per million tokens (3:1 input:output mix); the spread to the most expensive model is 386×.

This answer is computed from live data on every load — per-provider daily medians, per single GPU per hour, pricing tiers never mixed (see methodology). Quote it with attribution: “Data: Price of Compute — priceofcompute.com” (cite).