Report · updated 2026-08-28 · always live

The same AI model can cost you up to 14× more — depending on where you run it.

An analysis of 1,502 models from 86 vendors across 38 hosting providers, priced in USD per 1M tokens (blended 3:1 input:output). Every figure is computed live from the Verna One model index — so this page updates itself.

Key findings

The same model, 14× the price

When a model has open weights, several providers host the exact same weights and price them independently. That is not a rounding difference — it is a different budget:

ModelPrice spreadHostsCheapest
DeepSeek V3.2 14.0× 16 StreamLake — $0.241/1M
Llama 3.2 1B Instruct 11.3× 3 DeepInfra — $0.0063/1M
Qwen2.5 Coder 32B Instruct 9.0× 2 DeepInfra — $0.083/1M
Llama 3.1 8B Instruct 8.8× 6 DeepInfra — $0.025/1M
Gemma 4 31B 7.6× 24 OpenInference — $0.147/1M
MythoMax 13B 7.5× 8 OpenRouter (default route) — $0.060/1M
Nano Banana Pro (Gemini 3 Pro Image) 7.0× 2 OpenRouter (default route) — $4.50/1M
gpt-oss-120b 6.9× 24 CoreWeave — $0.065/1M

If you deployed on your provider's default and never re-checked, you are very likely paying multiples of the going rate for the identical output.

Who's actually cheapest

Across every model where we can compare, the cheapest option is:

ProviderModels it's cheapest for
OpenRouter (default route)91
OpenAI43
DeepInfra35
Alibaba23
Anthropic16
Mistral12

The pattern: closed models are usually cheapest from the vendor itself; open models are cheapest from specialist hosts or an aggregator route. There is no single "cheapest provider" — it is per model.

Best value right now

Quality per dollar (capability index ÷ price), from the live index:

ModelTierCheapest
Phi 4midDeepInfra — $0.087/1M
MiMo-V2.5frontierXiaomi — $0.175/1M
Granite 4.0 MicroeconomyOpenRouter (default route) — $0.041/1M
GPT-5.6 LunafrontierOpenAI — $0.225/1M
Nova Lite 1.0midOpenRouter (default route) — $0.105/1M
Mistral Small 3economyOpenRouter (default route) — $0.058/1M

What this means

Check your own models

The full, live data — every provider for every model — is free, no signup.

Open the model price index →

Methodology

Prices are normalised to USD per 1M tokens, blended 3:1 input:output, sourced from provider APIs and pricing pages and verified with dates. "Spread" is the ratio of the most- to least-expensive host of the same model, after reducing each provider to its own best price. Figures are as of 2026-08-28 and regenerate with the index. Full data: verna.one/models.

Frequently asked

How much can the same AI model vary in price?

For open-weight models hosted by several providers, the cost per million tokens can differ by as much as 14× for the identical model. Across the 137 multi-hosted models in the index, 65 vary by more than 2×, 12 by more than 5×, and 2 by more than 10×.

Which provider is cheapest?

There is no single cheapest provider — it is per model. Closed models are usually cheapest from the vendor itself; open-weight models are cheapest from specialist hosts or via an aggregator route. The live index shows the cheapest option for each model.

Is the data free and current?

Yes. The full index is free at verna.one/models, covers 1,502 models across 86 vendors and 38 hosting providers, and this report is regenerated from it — figures shown are as of 2026-08-28.