LLM API Pricing Comparison
Official API prices for 36 text models from OpenAI, Claude, Gemini, DeepSeek, Qwen, Kimi, GLM, MiniMax, MiMo and StepFun, per 1 million tokens. Input, output and cached input prices come from each provider's public pricing page.
Prices checked Oct 10, 2026 · Exchange rates Oct 10, 2026
Models
36
Model families
10
Unit
Per 1M tokens
Prices checked
Oct 10, 2026
Official API prices by model
Prices are in US dollars per 1 million tokens unless marked CN¥. CN¥ prices are published in Chinese yuan by the provider; the ≈ USD value uses 6.7065 CNY per USD (Oct 10, 2026).
| Model | Family | Input | Output | Cached input | Context window |
|---|---|---|---|---|---|
| GPT-5.5 | OpenAI | $5 | $30 | $0.50 | — |
| GPT-5.4 | OpenAI | $2.50 | $15 | $0.25 | — |
| GPT-5.4 Mini | OpenAI | $0.75 | $4.50 | $0.075 | — |
| GPT-5.4 Nano | OpenAI | $0.20 | $1.25 | $0.02 | — |
| GPT-6 Astra | OpenAI | $10 | $50 | $1 | 1.05M context · 128K output |
| GPT-6 Luna | OpenAI | $0.10 | $0.50 | $0.01 | 1.05M context · 128K output |
| GPT-6 Sol | OpenAI | $2 | $10 | $0.20 | 1.05M context · 128K output |
| Claude Opus 5 | Claude | $5 | $25 | $0.50 | 1M context · 128K max output |
| Claude Fable 5 | Claude | $10 | $50 | $1 | — |
| Claude Opus 4.8 | Claude | $5 | $25 | $0.50 | — |
| Claude Opus 4.7 | Claude | $5 | $25 | $0.50 | — |
| Claude Opus 4.6 | Claude | $5 | $25 | $0.50 | — |
| Claude Sonnet 5 | Claude | $2 | $10 | $0.20 | — |
| Claude Sonnet 4.6 | Claude | $3 | $15 | $0.30 | — |
| Claude Haiku 4.5 | Claude | $1 | $5 | $0.10 | — |
| Gemini 3.1 Pro Preview | Gemini | $2 | $12 | $0.20 | — |
| Gemini 3.1 Flash-Lite | Gemini | $0.25 | $1.50 | $0.025 | — |
| Gemini 3 Flash Preview | Gemini | $0.50 | $3 | $0.05 | — |
| DeepSeek V4 Flash | DeepSeek | $0.14 | $0.28 | $0.0028 | 1M |
| DeepSeek V4 Pro | DeepSeek | $0.435 | $0.87 | $0.003625 | 1M |
| Qwen3.7 Max | Qwen | CN¥12 (≈ $1.79) | CN¥36 (≈ $5.37) | — | — |
| Qwen3.7 Plus | Qwen | CN¥2 (≈ $0.30) | CN¥8 (≈ $1.19) | — | — |
| Qwen3.6 Plus | Qwen | CN¥2 (≈ $0.30) | CN¥12 (≈ $1.79) | — | — |
| Qwen3.5 Plus | Qwen | CN¥0.80 (≈ $0.12) | CN¥4.80 (≈ $0.72) | — | — |
| Qwen3 Coder Plus | Qwen | CN¥4 (≈ $0.60) | CN¥16 (≈ $2.39) | — | — |
| Kimi K3 | Kimi | CN¥20 (≈ $2.98) | CN¥100 (≈ $14.91) | CN¥2 (≈ $0.30) | 1M |
| Kimi K2.6 | Kimi | CN¥6.50 (≈ $0.97) | CN¥27 (≈ $4.03) | CN¥1.10 (≈ $0.16) | 256K |
| Kimi K2.5 | Kimi | CN¥4 (≈ $0.60) | CN¥21 (≈ $3.13) | CN¥0.70 (≈ $0.10) | 256K |
| GLM-5.1 | GLM | CN¥6 (≈ $0.89) | CN¥24 (≈ $3.58) | CN¥1.30 (≈ $0.19) | — |
| GLM-5 | GLM | CN¥4 (≈ $0.60) | CN¥18 (≈ $2.68) | CN¥1 (≈ $0.15) | 200K |
| MiniMax M3 | MiniMax | $0.30 | $1.20 | $0.06 | 1M |
| MiniMax M2.7 | MiniMax | $0.30 | $1.20 | $0.06 | 204.8K |
| MiniMax M2.5 | MiniMax | $0.30 | $1.20 | $0.03 | 204.8K |
| MiMo V2.5 Pro | MiMo | $0.435 | $0.87 | $0.0036 | 1M |
| MiMo V2.5 | MiMo | $0.14 | $0.28 | $0.0028 | 1M |
| Step 3.7 Flash | StepFun | CN¥1.35 (≈ $0.20) | CN¥8.10 (≈ $1.21) | CN¥0.27 (≈ $0.04) | — |
How to read these prices
- Input is the price for tokens you send to the model, including the prompt, system message and conversation history.
- Output is the price for tokens the model generates. Reasoning tokens are usually billed as output.
- Cached input is the discounted price for prompt tokens served from the provider's prompt cache.
- Some providers charge different rates for very long prompts, batch jobs or priority processing. Open a model page for every provider and plan we track, and check the provider's pricing page before you build on it.
BiWu.Ai lists public API prices from official pricing pages for reference. We do not sell API access, and prices can change without notice.