Runyard / LLM API Index · GPU Index →

LLM API Pricing Index

Input and output price per million tokens for 440 models from 59 vendors, in one schema. 418 are paid, 22 have a free tier. Blended price assumes 3 input tokens per output token, the ratio most chat and agent work lands near.

Last collected 2026-09-12 · 7 snapshots

Models priced
44059 vendors
Median paid model
$0.850/1M blended
Frontier tier
$20.00/1M blended
Cheapest paid
$0.022Mistral Nemo
Flagships, blended $/1M tokens
3 input : 1 output · as listed on OpenRouter
Claude Fable 5.1$20.00GPT-6 Astra$20.00Grok 4.6$3.00Kimi K3$5.31DeepSeek V4.1 Flash$0.26GLM 5.3 Flash$0.24Gemini 3.8 Flash$1.50

The spread is the story: the frontier tier costs 13× the cheapest flagship-class model on this chart, for work that often does not need it.

Every model

ContextCache read7D
Aion-3.0-MiniAion Labs · aion-labs/aion-3.0-mini131K$0.700$1.40$0.875$0.180
Aion-2.0Aion Labs · aion-labs/aion-2.0131K$0.800$1.60$1.00$0.200
Aion-RP 1.0 (8B)Aion Labs · aion-labs/aion-rp-llama-3.1-8b33K$0.800$1.60$1.00
Aion-3.0Aion Labs · aion-labs/aion-3.0131K$3.00$6.00$3.75$0.750

4 of 4 models (free tiers hidden). Blended = 3 input tokens per output token. Prices as listed on OpenRouter on the date above.

Vendors

VendorModelsCheapest paid, blended
OpenAI93$0.055
Alibaba Qwen53$0.055
Google43$0.063
Anthropic27$0.500
Mistral25$0.022
DeepSeek18$0.050
Z.ai (Zhipu)17$0.119
NVIDIA10$0.087
MiniMax9$0.425
Moonshot AI8$0.900
Meta8$0.058
Meta7$0.125
Tencent7$0.077
xAI7$1.25
inclusionAI6$0.032
Bytedance Seed6$0.131
Thinking Machines6$0.637
~Openai5$0.450
Cohere5$0.066
Amazon5$0.061
Perplexity5$1.00
Sakana AI4$1.71
Poolside4$0.075
Aion Labs4$0.875
~Anthropic4$2.00
Thedrummer3$0.350
Nous Research3$0.700
Sao10k3$0.043
Inference Net2$0.060
Inception2$0.068
Nex Agi2
Ibm Granite2$0.041
~Z Ai2$0.119
Upstage2$0.158
Kwaipilot2$0.525
Stepfun2$0.150
~Google2$1.50
Xiaomi2$0.175
Rekaai2$0.100
Relace2$0.950
Morph2$0.900
Microsoft2$0.087
Dots Studio1
Liquid AI1
~Deepseek1$0.040
Meituan1$0.525
~X Ai1$3.00
Perceptron1$0.487
~Moonshotai1$5.11
Arcee Ai1$0.388
Openrouter1
Writer1$1.95
Bytedance1$0.125
Cognitivecomputations1$0.375
Baidu1$0.627
Anthracite Org1$3.13
Mancer1$0.487
Undi951$0.425
Gryphe1$0.060

How this is built

Prices are read from OpenRouter's public model list, which routes to each vendor and publishes the per-token price it charges — for nearly every model, the vendor's own list price passed through. It is the one place all vendors appear in one schema, which is why it is the source. Prices are shown as listed there; a vendor's direct price can differ, especially with volume discounts or provisioned throughput, and OpenRouter's own listing is the reference for what is shown.

Every collector run is stored as a snapshot and committed to the repository, so price changes are diffs anyone can inspect. Nothing is backfilled. The 7-day column compares a model against itself a week earlier and is blank until that run exists.

Blended weights input at 3 and output at 1. Change the ratio and the ranking changes — an output-heavy workload makes the frontier tier relatively more expensive, since output tokens cost five times input on most frontier models.

What the same tokens cost on hardware you rent or own.
Local cost per million tokens →