Stack Spend
Live dataset · Updated daily

The LLM API Pricing Index

List prices per 1M tokens across the major model providers and inference hosts, next to each model’s coding-benchmark score — so you can spot equal-or-better quality for less. Synced daily from public sources.

Estimate your bill

Prices updated 7 October 2026

At a glance

LLM API pricing, compared across providers and updated daily. This page lists list prices in USD per 1 million tokens for 12 providers — OpenAI, Anthropic, Google (Gemini), xAI, DeepSeek, Mistral — and inference hosts including Groq, Together AI, Fireworks AI and DeepInfra, alongside each model’s coding-benchmark score (best of Aider polyglot and SWE-bench Verified). Prices come from the LiteLLM price dataset; benchmarks from the Aider polyglot leaderboard and the Epoch AI Benchmarking Hub. For coding value, deepseek-ai/DeepSeek-V3.2 scores 74.2% at $0.38 per 1M output tokens.

983

models tracked across 12 providers, updated 7 October 2026

83.5%

top coding score — claude-opus-4-7 (Anthropic)

$0.38 /1M out

best coding value — deepseek-ai/DeepSeek-V3.2 at 74.2%

gpt-5.5OpenAI$5.00$30.001.1M80.6%
gpt-5.4OpenAI$2.50$15.001.1M76.9%
gpt-5.4-miniOpenAI$0.75$4.50272K—
o3OpenAI$2.00$8.00200K76.9%
claude-opus-4-8Anthropic$5.00$25.001M—
claude-sonnet-4-5SunsettingAnthropic$3.00$15.001M71.3%
claude-haiku-4-5Anthropic$1.00$5.00200K—
gemini-2.5-proGoogle (Gemini)$1.25$10.001.0M83.1%
gemini-2.5-flashGoogle (Gemini)$0.30$2.501.0M55.1%
deepseek-v3.2DeepSeek$0.28$0.40164K74.2%
deepseek-r1DeepSeek$0.55$2.1966K56.9%
mistral-large-latestMistral$0.50$1.50262K—
claude-opus-4-7Anthropic$5.00$25.001M83.5%
anthropic/claude-opus-4-7DeepInfra$5.00$25.001M83.5%
anthropic/claude-opus-4-7Perplexity$5.00$25.00—83.5%
google/gemini-2.5-proDeepInfra$1.25$10.001M83.1%
gemini-2.5-pro-preview-ttsGoogle (Gemini)$1.00$20.008K83.1%
openai/gpt-5.5Perplexity$5.00$30.00—80.6%
google/gemini-3.5-flashDeepInfra$1.50$9.001M79.3%
gemini-3.5-flashGoogle (Gemini)$1.50$9.001.0M79.3%
google/gemini-3.5-flashPerplexity$1.50$9.00—79.3%
claude-opus-4-6Anthropic$5.00$25.001M78.7%
anthropic/claude-opus-4-6Perplexity$5.00$25.00—78.7%
zai-org/GLM-5.2DeepInfra$0.75$2.401.0M78.7%
glm-5-2Mistral$1.40$4.401.0M78.7%
perplexity/glm-5.2Perplexity$1.40$4.40—78.7%
zai-org/GLM-5.2Together AI$1.40$4.401.0M78.7%
deepseek-ai/DeepSeek-V4-ProDeepInfra$1.30$2.601.0M77.6%
deepseek-ai/DeepSeek-V4-Pro-0813DeepInfra$1.30$2.601.0M77.6%
deepseek-v4-proDeepSeek$1.32$3.961M77.6%
Qwen/Qwen3.7-MaxTogether AI$1.50$4.501M77.3%
Qwen/Qwen3.7-MaxDeepInfra$2.50$7.50256K77.3%
openai/gpt-5.4Perplexity$2.50$15.00—76.9%
moonshotai/Kimi-K2.6DeepInfra$0.75$3.50262K76.7%
kimi-k2.6Moonshot AI$0.95$4.00262K76.7%
claude-opus-4-5Anthropic$5.00$25.00200K76.7%
anthropic/claude-opus-4-5Perplexity$5.00$25.00—76.7%
google/gemini-3.1-proDeepInfra$2.00$12.001M75.6%
gemini-3.1-pro-previewGoogle (Gemini)$2.00$12.001.0M75.6%
gemini-3.1-pro-preview-customtoolsGoogle (Gemini)$2.00$12.001.0M75.6%
google/gemini-3.1-pro-previewPerplexity$2.00$12.00—75.6%
gemini-3-flash-previewGoogle (Gemini)$0.50$3.001.0M75.4%
google/gemini-3-flash-previewPerplexity$0.50$3.00—75.4%
claude-sonnet-4-6Anthropic$3.00$15.001M75.2%
anthropic/claude-sonnet-4-6DeepInfra$3.00$15.001M75.2%
anthropic/claude-sonnet-4-6Perplexity$3.00$15.00—75.2%
gpt-5.3-codexSunsettingOpenAI$1.75$14.00272K74.8%
deepseek-ai/DeepSeek-V3.2DeepInfra$0.26$0.38164K74.2%
accounts/fireworks/models/deepseek-v3p2Fireworks AI$0.56$1.68164K74.2%
zai-org/GLM-5.1DeepInfra$1.05$3.50203K74.2%

Showing 50 of 777 models · prices in USD per 1M tokens · benchmarks by base model: coding via Aider polyglot & SWE-bench Verified; reasoning & math via Epoch AI.

Get price-change alerts

One email when a model’s price changes, a new model launches, or a model you might rely on gets a deprecation date. No schedule, no newsletter — it only sends on days something actually changed.

Unsubscribe with one click, any time.

Methodology & sources

The StackSpend LLM Pricing Index is refreshed daily. Prices are provider list prices in USD per 1 million tokens, synced from the community-maintained LiteLLM price dataset. Coding scores are the best of the Aider polyglot benchmark and SWE-bench Verified (via the Epoch AI Benchmarking Hub, CC BY) and are attached to the underlying base model, so the identical open-weight model served by different hosts shares one score. StackSpend uses this same data to track and right-size your own AI spend. For what these models are and how the families relate, see the LLM model glossary.

How often is this LLM pricing data updated?
Daily. Prices are synced from the community-maintained LiteLLM price dataset; coding scores are the best of the Aider polyglot benchmark and SWE-bench Verified (via Epoch AI); this page was last refreshed on 7 October 2026. All prices are list prices in USD per 1M tokens.
Which model gives the best coding performance for the price?
Among models benchmarked for coding, deepseek-ai/DeepSeek-V3.2 (DeepInfra) offers the strongest coding score relative to its output price — a 74.2% coding-benchmark score (best of Aider polyglot and SWE-bench Verified) at $0.38 per 1M output tokens.
What is the highest-scoring model for coding?
claude-opus-4-7 (Anthropic) currently leads the coding benchmark at 83.5%, priced at $5 input / $25 output per 1M tokens.
Why do the same open models cost different amounts?
Open-weight models (like Llama or DeepSeek) are served by multiple inference hosts — Groq, Together AI, Fireworks, DeepInfra and others — at different prices for the identical weights. Comparing hosts for the same model is often the fastest saving.

Use this data

The Index is free to use under CC BY 4.0 — download it or pull it live, and cite the StackSpend LLM Pricing Index with a link back to this page.

Cite as: StackSpend LLM Pricing Index — https://www.stackspend.app/resources/llm-api-pricing (updated 7 October 2026).

Track what you actually spend on these models.

StackSpend connects to your AI and cloud providers read-only, normalises every model into one daily signal, and flags when a model or deploy pushes your spend off-baseline — with cheaper equal-quality alternatives from this very dataset.