AI APIs · Price comparison

AI API pricing comparison

Compare public list prices across 11 current AI models from five providers. Every number below is priced per 1 million tokens and links back to a source. Use the calculator for your own input/output mix — the cheapest input rate is not always the cheapest workload.

What is the cheapest AI API here?

Cheapest input tokens

Qwen3.5 Flash

$0.065 per 1M input tokens

Cheapest output tokens

Qwen3.5 Flash

$0.26 per 1M output tokens

This is a rate-card answer, not a quality ranking. A model that needs more output tokens, retries, or human correction can cost more per completed task even when its listed token price is lower.

LLM API price comparison per 1M tokens

ModelProviderInput / 1MOutput / 1MCached inputContext
Qwen3.5 FlashAlibaba$0.065$0.261M
Gemini 3.5 Flash-LiteGoogle$0.30$2.50Not stated
Qwen3.5 PlusAlibaba$0.30$1.801M
Qwen3.7 PlusAlibaba$0.32$1.281M
GPT-5.6 LunaOpenAI$1.00$6.001M
Qwen3.7 MaxAlibaba$1.48$4.431M
Gemini 3.6 FlashGoogle$1.50$7.50Not stated
GPT-5.6 TerraOpenAI$2.50$15.001M
Kimi K3Moonshot AI$3.00$15.00$0.301M
GPT-5.6 SolOpenAI$5.00$30.001M
Claude Opus 5Anthropic$5.00$25.00Not stated

Sorted by input price. Prices exclude batch discounts, provider promotions, enterprise agreements and tool-call charges. “Not stated” means the cited source did not publish a context window; it is not an estimate.

AI token cost calculator

Enter the same workload for every model. The calculator combines input and output cost and marks the lowest list-price total for that token mix.

Cost calculator

Enter your monthly token volume to see what each tier would cost.

TierInput costOutput costMonthly total
Qwen3.5 Flash$0.065$0.052$0.117cheapest
Gemini 3.5 Flash-Lite$0.30$0.50$0.80
Qwen3.5 Plus$0.30$0.36$0.66
Qwen3.7 Plus$0.32$0.256$0.576
Luna$1.00$1.20$2.20
Qwen3.7 Max$1.48$0.885$2.36
Gemini 3.6 Flash$1.50$1.50$3.00
Terra$2.50$3.00$5.50
Kimi K3$3.00$3.00$6.00
Sol$5.00$6.00$11.00
Opus 5$5.00$5.00$10.00

Estimates only — list prices, before any caching discount, batch pricing, or enterprise agreement.

How to compare AI API pricing without fooling yourself

Separate input and output

Output is often several times more expensive. A chat app and a document classifier can rank models differently even at the same total token volume.

Measure tokens per accepted result

Reasoning verbosity, retries and rejected answers change the real bill. Benchmark the task you actually run, not only the vendor's per-token rate.

Recheck the rate card

Vendors change prices and promotions. Each row carries sourced data, and the page shows when this comparison was last verified.

Need detail on one family? Start with the GPT-5.6 price breakdown or the Kimi K3 cost analysis.

Sources

Figures on this page last checked against these sources on 2026-07-25. Vendors change pricing and specs without notice — if a number here disagrees with the vendor's own page, trust the vendor.