AI APIs · Price comparison

AI API pricing comparison

Compare official vendor list prices separately from third-party channel prices. Every number below is priced per 1 million tokens, includes its applicable prompt range, and links back to a source.

What is the cheapest AI API here?

Cheapest input tokens

GPT-5.6 Luna

$0.20 per 1M input tokens

Official list price for < 272K.

Cheapest output tokens

GPT-5.6 Luna

$1.20 per 1M output tokens

Official list price for < 272K.

This is a rate-card answer, not a quality ranking. A model that needs more output tokens, retries, or human correction can cost more per completed task even when its listed token price is lower.

LLM API price comparison per 1M tokens

This first table is official list pricing only. Every context-dependent rate is expanded into its own row. OpenRouter and other retail/channel prices are kept in the separate table below and never determine the “cheapest official API” headline.

ModelProviderInput / 1MOutput / 1MCached inputApplicable promptContext
GPT-5.6 LunaOpenAI$0.20$1.20< 272K1.05M
GPT-5.6 LunaOpenAI$0.40$1.80≥ 272K1.05M
Gemini 3.5 Flash-LiteGoogle$0.30$2.50All promptsNot stated
Qwen3.5 PlusAlibaba$0.40$2.40≤ 256K1M
Qwen3.5 PlusAlibaba$0.50$3.00> 256K1M
Gemini 3.7 FlashGoogle$0.75$3.75All promptsNot stated
Claude Haiku 4.5Anthropic$1.00$5.00$0.10All promptsNot stated
Gemini 3.6 FlashGoogle$1.50$7.50All promptsNot stated
Grok 4.6xAI$2.00$6.00< Not statedNot stated
Grok 4.6xAI$4.00$12.00All promptsNot stated
GPT-5.6 TerraOpenAI$2.00$12.00< 272K1.05M
GPT-5.6 TerraOpenAI$4.00$18.00≥ 272K1.05M
Claude Sonnet 5Anthropic$2.00$10.00$0.20All prompts1M
GPT-5.6 SolOpenAI$5.00$30.00< 272K1.05M
GPT-5.6 SolOpenAI$10.00$45.00≥ 272K1.05M
Claude Opus 5Anthropic$5.00$25.00< Not statedNot stated
Claude Opus 5Anthropic$10.00$50.00All promptsNot stated

Models are sorted by base input price. Prices exclude batch discounts, promotions, enterprise agreements and tool-call charges. “Not stated” means the cited source did not publish a context window; it is not an estimate.

OpenRouter channel pricing

These are third-party channel rates, not vendor canonical list prices. They can include retail discounts or promotions and may differ from buying from the vendor.

ModelInput / 1MOutput / 1MCached inputApplicable promptContext
Qwen3.7 Flash$0.03$0.13$0.006< 32K1M
Qwen3.7 Flash$0.10$0.40$0.02≥ 32K – < 256K1M
Qwen3.7 Flash$0.20$0.80$0.04≥ 256K1M
Qwen3.5 Flash$0.065$0.26All prompts1M
Qwen3.6 Flash$0.1875$1.125< 256K1M
Qwen3.6 Flash$0.75$3.00≥ 256K1M
Gemini 3.1 Flash Lite$0.25$1.50$0.025All prompts1.048576M
Qwen3.5 Plus$0.30$1.80< 256K1M
Qwen3.5 Plus$0.375$2.25≥ 256K1M
Qwen3.7 Plus$0.32$1.28$0.064< 256K1M
Qwen3.7 Plus$0.96$3.84$0.192≥ 256K1M
Qwen3.6 Plus$0.325$1.95< 256K1M
Qwen3.6 Plus$1.30$3.90≥ 256K1M
Kimi K2.7 Code$0.71$3.50$0.15All prompts262.144K
Qwen3.6 Max Preview$1.027$6.162< 128K262.144K
Qwen3.6 Max Preview$1.58$9.48≥ 128K262.144K
Qwen3.7 Max$1.475$4.425All prompts1M
GPT-5.6 Sol$2.50$15.00$0.25All prompts1.05M
Kimi K3$3.00$15.00$0.30All prompts1M

AI token cost calculator

Enter the same workload for every model. The calculator combines input and output cost and marks the lowest list-price total for that token mix.

Cost calculator

Enter monthly volume and the typical prompt size. The applicable context tier is selected automatically.

TierInput costOutput costMonthly total
Luna$0.20$0.24$0.44cheapest
Gemini 3.5 Flash-Lite$0.30$0.50$0.80
Qwen3.5 Plus (Alibaba international, ≤256K)$0.40$0.48$0.88
Gemini 3.7 Flash (introductory, through 2026-12-31)$0.75$0.75$1.50
Haiku 4.5$1.00$1.00$2.00
Gemini 3.6 Flash$1.50$1.50$3.00
Grok 4.6$2.00$1.20$3.20
Terra$2.00$2.40$4.40
Sonnet 5 (introductory, through 2026-08-31)$2.00$2.00$4.00
Sol$5.00$6.00$11.00
Opus 5$5.00$5.00$10.00

Estimates only — list prices, before any caching discount, batch pricing, or enterprise agreement.

How to compare AI API pricing without fooling yourself

Separate input and output

Output is often several times more expensive. A chat app and a document classifier can rank models differently even at the same total token volume.

Measure tokens per accepted result

Reasoning verbosity, retries and rejected answers change the real bill. Benchmark the task you actually run, not only the vendor's per-token rate.

Recheck the rate card

Vendors change prices and promotions. Each row carries sourced data, and the page shows when this comparison was last verified.

Need detail on one family? Start with the GPT-5.6 price breakdown or the Kimi K3 cost analysis.

Sources

Figures on this page last checked against these sources on 2026-08-17. Vendors change pricing and specs without notice — if a number here disagrees with the vendor's own page, trust the vendor.