← the bench index · API pricing

LLM API pricing comparison

What 29 live models cost on their makers' own APIs, in US dollars per million tokens, from 11 labs. These are the published list prices in the models.dev catalog as of September 24, 2026, refreshed daily. Clouds and resellers may charge differently. The AW Index column is each model's standing on the bench index, which combines the public leaderboards.

Get the next launch alert, free →

Models on the bench index

ModelInputOutputCached inputContextAW Index
Claude Opus 5.5Anthropic · claude-opus-5-5$4$20$0.201M100
Fable 5.1Anthropic · claude-fable-5-1$10$50$0.251M95
GPT-6 AstraOpenAI · gpt-6-astra$10$50$11.05M88.5
Claude Opus 5Anthropic · claude-opus-5$5$25$0.501M82.7
Fable 5Anthropic · claude-fable-5$10$50$11M81.4
ChatGPT-5.6OpenAI · gpt-5.6$4$20$0.401.05M70.7
Grok 4.7xAI · grok-4.7$2$6$0.50500K69.7
Kimi K3Moonshot · kimi-k3$3$15$0.301.05M60.4
GPT-5.5OpenAI · gpt-5.5$5$30$0.501.05M52.4
Qwen 3.8 MaxAlibaba · qwen3.8-max$2$6$0.251M50.2
Claude Opus 4.8Anthropic · claude-opus-4-8$5$25$0.501M47.8
Grok 4.5xAI · grok-4.5$2$6$0.30500K47.4
Claude Sonnet 5Anthropic · claude-sonnet-5$2$10$0.201M43.5
Muse Spark 1.1Meta · muse-spark-1.1$1.25$4.25$0.151.05M37.7
DeepSeek V4 ProDeepSeek · deepseek-v4-pro$0.435$0.87$0.0036251M37.1
GLM-5.2Zhipu AI · glm-5.2$1.40$4.40$0.261M31.8
Gemini 3.5 FlashGoogle · gemini-3.5-flash$1.50$9$0.151.05M27.3
MiniMax M3MiniMax · MiniMax-M3$0.30$1.20$0.061M4.3

More live models

ModelInputOutputCached inputContext
Grok 4.6xAI · grok-4.6$2$6$0.50500K
GLM-5.3Zhipu AI · glm-5.3$1.40$4.40$0.261M
Muse Spark 1.2Meta · muse-spark-1.2$1.25$4.25$0.151.05M
Muse Spark 1.3Meta · muse-spark-1.3$1.25$4.25$0.151.05M
Gemini 3.6 FlashGoogle · gemini-3.6-flash$0.75$3.75$0.0751.05M
Gemini 3.7 FlashGoogle · gemini-3.7-flash$0.75$3.75$0.0751.05M
Gemini 3.8 FlashGoogle · gemini-3.8-flash$0.75$3.75$0.0751.05M
Gemini 3.5 Flash-LiteGoogle · gemini-3.5-flash-lite$0.30$2.50$0.031.05M
GLM-5.3 FlashZhipu AI · glm-5.3-flash$0.15$0.50$0.031M
Qwen 3.8 FlashAlibaba · qwen3.8-flash$0.15$0.47$0.0161M
MiMo-V2.5Xiaomi · mimo-v2.5$0.14$0.28$0.00281.05M

† Long prompts cost more: GPT-6 Astra $20 / $75 over 272K tokens; ChatGPT-5.6 $8 / $30 over 272K tokens; Grok 4.7 $4 / $12 over 200K tokens; GPT-5.5 $10 / $45 over 272K tokens; Grok 4.5 $4 / $12 over 200K tokens; MiniMax M3 $0.60 / $2.40 over 512K tokens; Grok 4.6 $4 / $12 over 200K tokens.

Quick answers

Which LLM API is the cheapest?

Of the 29 live models priced here, MiMo-V2.5 is the cheapest on both input and output: $0.14 input / $0.28 output per million tokens.

Which benchmarked model is the cheapest?

Among the 18 priced models scored on the ArtificialWatch bench index, the lowest output price is DeepSeek V4 Pro, at $0.435 input / $0.87 output per million tokens. It stands at 37.1 on the index, where the top model is at 100.

Which LLM API is the most expensive?

The highest output price of the 29 priced here is $50 per million tokens: Fable 5.1, GPT-6 Astra and Fable 5 all charge it.

Where do these prices come from?

Each model maker's own API listing in the models.dev catalog, as of 2026-09-24, refreshed daily. A price is shown only for the catalog id that is the model's exact name, so a cheaper sibling tier never appears under a flagship's name. Clouds and resellers may charge differently; each model's page lists the providers that carry it.

Price is one axis. For scores side by side, see the head-to-head comparisons and the bench index. To hear the minute a new model goes live, watch free.