Runyard / LLM API Index · GPU Index →

LLM API Pricing Index

Input and output price per million tokens for 440 models from 59 vendors, in one schema. 418 are paid, 22 have a free tier. Blended price assumes 3 input tokens per output token, the ratio most chat and agent work lands near.

Last collected 2026-09-12 · 7 snapshots

Models priced
44059 vendors
Median paid model
$0.850/1M blended
Frontier tier
$20.00/1M blended
Cheapest paid
$0.022Mistral Nemo
Flagships, blended $/1M tokens
3 input : 1 output · as listed on OpenRouter
Claude Fable 5.1$20.00GPT-6 Astra$20.00Grok 4.6$3.00Kimi K3$5.31DeepSeek V4.1 Flash$0.26GLM 5.3 Flash$0.24Gemini 3.8 Flash$1.50

The spread is the story: the frontier tier costs 13× the cheapest flagship-class model on this chart, for work that often does not need it.

Every model

ContextCache read7D
Gemma 3 4BGoogle · google/gemma-3-4b-it131K$0.050$0.100$0.063
Gemma 3 12BGoogle · google/gemma-3-12b-it131K$0.050$0.150$0.075
Gemma 4 26B A4B Google · google/gemma-4-26b-a4b-it262K$0.042$0.220$0.086
Gemini 2.5 Flash Lite (batch)Google · google/gemini-2.5-flash-lite:batch1.0M$0.050$0.200$0.087$0.010
Gemma 4 31BGoogle · google/gemma-4-31b-it262K$0.090$0.340$0.152$0.050
Gemma 3 27BGoogle · google/gemma-3-27b-it131K$0.080$0.450$0.172$0.040
Gemini 2.5 Flash LiteGoogle · google/gemini-2.5-flash-lite1.0M$0.100$0.400$0.175$0.010
Gemini 3.1 Flash Lite (batch)Google · google/gemini-3.1-flash-lite:batch1.0M$0.125$0.750$0.281$0.013
Gemini 3.5 Flash Lite (batch)Google · google/gemini-3.5-flash-lite:batch1.0M$0.150$1.25$0.425$0.015
Gemini 2.5 Flash (batch)Google · google/gemini-2.5-flash:batch1.0M$0.150$1.25$0.425$0.030
Gemma 4 31B (batch)Google · google/gemma-4-31b-it:batch262K$0.390$0.970$0.535
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google · google/gemini-3.1-flash-lite-image66K$0.250$1.50$0.563
Gemini 3.1 Flash LiteGoogle · google/gemini-3.1-flash-lite1.0M$0.250$1.50$0.563$0.025
Gemini 3.1 Flash Lite PreviewGoogle · google/gemini-3.1-flash-lite-preview1.0M$0.250$1.50$0.563$0.025
Gemini 3 Flash Preview (batch)Google · google/gemini-3-flash-preview:batch1.0M$0.250$1.50$0.563
Gemma 2 27BGoogle · google/gemma-2-27b-it8K$0.650$0.650$0.650
Gemini 3.8 Flash (batch)Google · google/gemini-3.8-flash:batch1.0M$0.375$1.88$0.750$0.037
Gemini 3.7 Flash (batch)Google · google/gemini-3.7-flash:batch1.0M$0.375$1.88$0.750$0.037
Gemini 3.6 Flash (batch)Google · google/gemini-3.6-flash:batch1.0M$0.375$1.88$0.750$0.037
Gemini 3.5 Flash LiteGoogle · google/gemini-3.5-flash-lite1.0M$0.300$2.50$0.850$0.030
Nano Banana (Gemini 2.5 Flash Image)Google · google/gemini-2.5-flash-image33K$0.300$2.50$0.850$0.030
Gemini 2.5 FlashGoogle · google/gemini-2.5-flash1.0M$0.300$2.50$0.850$0.030
Nano Banana 2 (Gemini 3.1 Flash Image)Google · google/gemini-3.1-flash-image131K$0.500$3.00$1.13
Nano Banana 2 (Gemini 3.1 Flash Image Preview)Google · google/gemini-3.1-flash-image-preview66K$0.500$3.00$1.13
Gemini 3 Flash PreviewGoogle · google/gemini-3-flash-preview1.0M$0.500$3.00$1.13$0.050
Gemini 3.8 FlashGoogle · google/gemini-3.8-flash1.0M$0.750$3.75$1.50$0.075
Gemini 3.7 FlashGoogle · google/gemini-3.7-flash1.0M$0.750$3.75$1.50$0.075
Gemini 3.6 FlashGoogle · google/gemini-3.6-flash1.0M$0.750$3.75$1.50$0.075
Gemini 3.5 Flash (batch)Google · google/gemini-3.5-flash:batch1.0M$0.750$4.50$1.69$0.075
Gemini 2.5 Pro (batch)Google · google/gemini-2.5-pro:batch1.0M$0.625$5.00$1.72$0.125
Gemini 3.1 Pro Preview (batch)Google · google/gemini-3.1-pro-preview:batch1.0M$1.00$6.00$2.25
Gemini 3.5 FlashGoogle · google/gemini-3.5-flash1.0M$1.50$9.00$3.38$0.150
Gemini 2.5 ProGoogle · google/gemini-2.5-pro1.0M$1.25$10.00$3.44$0.125
Gemini 2.5 Pro Preview 06-05Google · google/gemini-2.5-pro-preview1.0M$1.25$10.00$3.44$0.125
Gemini 2.5 Pro Preview 05-06Google · google/gemini-2.5-pro-preview-05-061.0M$1.25$10.00$3.44$0.125
Nano Banana Pro (Gemini 3 Pro Image)Google · google/gemini-3-pro-image131K$2.00$12.00$4.50$0.200
Gemini 3.1 Pro Preview Custom ToolsGoogle · google/gemini-3.1-pro-preview-customtools1.0M$2.00$12.00$4.50$0.200
Gemini 3.1 Pro PreviewGoogle · google/gemini-3.1-pro-preview1.0M$2.00$12.00$4.50$0.200
Nano Banana Pro (Gemini 3 Pro Image Preview)Google · google/gemini-3-pro-image-preview66K$2.00$12.00$4.50$0.200

39 of 39 models (free tiers hidden). Blended = 3 input tokens per output token. Prices as listed on OpenRouter on the date above.

Vendors

VendorModelsCheapest paid, blended
OpenAI93$0.055
Alibaba Qwen53$0.055
Google43$0.063
Anthropic27$0.500
Mistral25$0.022
DeepSeek18$0.050
Z.ai (Zhipu)17$0.119
NVIDIA10$0.087
MiniMax9$0.425
Moonshot AI8$0.900
Meta8$0.058
Meta7$0.125
Tencent7$0.077
xAI7$1.25
inclusionAI6$0.032
Bytedance Seed6$0.131
Thinking Machines6$0.637
~Openai5$0.450
Cohere5$0.066
Amazon5$0.061
Perplexity5$1.00
Sakana AI4$1.71
Poolside4$0.075
Aion Labs4$0.875
~Anthropic4$2.00
Thedrummer3$0.350
Nous Research3$0.700
Sao10k3$0.043
Inference Net2$0.060
Inception2$0.068
Nex Agi2
Ibm Granite2$0.041
~Z Ai2$0.119
Upstage2$0.158
Kwaipilot2$0.525
Stepfun2$0.150
~Google2$1.50
Xiaomi2$0.175
Rekaai2$0.100
Relace2$0.950
Morph2$0.900
Microsoft2$0.087
Dots Studio1
Liquid AI1
~Deepseek1$0.040
Meituan1$0.525
~X Ai1$3.00
Perceptron1$0.487
~Moonshotai1$5.11
Arcee Ai1$0.388
Openrouter1
Writer1$1.95
Bytedance1$0.125
Cognitivecomputations1$0.375
Baidu1$0.627
Anthracite Org1$3.13
Mancer1$0.487
Undi951$0.425
Gryphe1$0.060

How this is built

Prices are read from OpenRouter's public model list, which routes to each vendor and publishes the per-token price it charges — for nearly every model, the vendor's own list price passed through. It is the one place all vendors appear in one schema, which is why it is the source. Prices are shown as listed there; a vendor's direct price can differ, especially with volume discounts or provisioned throughput, and OpenRouter's own listing is the reference for what is shown.

Every collector run is stored as a snapshot and committed to the repository, so price changes are diffs anyone can inspect. Nothing is backfilled. The 7-day column compares a model against itself a week earlier and is blank until that run exists.

Blended weights input at 3 and output at 1. Change the ratio and the ranking changes — an output-heavy workload makes the frontier tier relatively more expensive, since output tokens cost five times input on most frontier models.

What the same tokens cost on hardware you rent or own.
Local cost per million tokens →