Find the Cheapest & Best-Fit LLM API
Real-time tracking of 63 providers, 540 models' pricing, capabilities & context window. Covering international and Chinese models, auto-updated daily at 04:00.
This site tracks public LLM API pricing and benchmarks with daily updates. Background, sources, and caveats live on the About page →
Best Value Models
Filters to models with AA Intelligence Index 40 or higher, then ranks by intelligence per dollar (AA score ÷ input price)
| # | Model | Provider | Input | Benchmark |
|---|---|---|---|---|
| 1 | Upstage: Solar Pro 4 | | $0.030 | AA Index 42 |
| 2 | DeepSeek: DeepSeek V4 Flash 0731 | | $0.060 | AA Index 52 |
| 3 | Z.ai: GLM 5.3 Flash Multimodal | | $0.075 | AA Index 57 |
| 4 | DeepSeek: DeepSeek V4 Flash 0423 | | $0.080 | AA Index 42 |
| 5 | Tencent: Hy3 | | $0.083 | AA Index 42 |
Strongest Models
Ranked by Artificial Analysis Intelligence Index, a cross-domain intelligence metric
| # | Model | Provider | Input | AA Index |
|---|---|---|---|---|
| 1 | Claude Opus 5 | | $5.00 | 63 |
| 2 | Anthropic: Claude Fable 5 | | $10.00 | 62 |
| 3 | OpenAI: GPT-5.6 Sol | | $2.00 | 61 |
| 4 | SpaceXAI: Grok 4.6 | | $2.00 | 61 |
| 5 | MoonshotAI: Kimi K3 | | $3.00 | 60 |
Cheapest Input Price
Cost in USD per million input tokens (excluding free models)
| # | Model | Provider | Input | Output |
|---|---|---|---|---|
| 1 | inclusionAI: Ling-2.6-flash | | $0.010 | $0.030 |
| 2 | IBM: Granite 4.0 Micro | | $0.017 | $0.112 |
| 3 | Mistral: Mistral Nemo | | $0.019 | $0.030 |
| 4 | Ling-3.0-flash | | $0.021 | $0.063 |
| 5 | OpenAI: GPT-5 Nano (batch) | | $0.025 | $0.200 |
Fastest Output
Ranked by Artificial Analysis measured output speed (tokens/sec)
| # | Model | Provider | t/s |
|---|---|---|---|
| 1 | Inception: Mercury 2 | | 963 |
| 2 | Google: Gemini 2.5 Flash Lite | | 395 |
| 3 | Ling-3.0-flash | | 383 |
| 4 | Google: Gemini 3.5 Flash Lite | | 355 |
| 5 | Google: Gemini 3.7 Flash | | 338 |
Free Models (56)
Models with $0 input price. Some may still charge for output — click to view full pricing details.
| Model | Provider | Output $/M | Context |
|---|---|---|---|
| Arcee AI: Trinity Large Thinking (free) | | Free | 262K |
| Baidu Qianfan: CoBuddy (free) | | Free | 131K |
| Baidu: Qianfan-OCR-Fast (free) | | Free | 66K |
| Cohere: North Mini Code (free) | | Free | 256K |
| DeepSeek: DeepSeek V4 Flash (free) | | Free | 1.05M |
Top Providers
🇺🇸 OpenAI
- Models
- 108
- Cheapest Input
- $0.025 /M tokens
🇨🇳 Qwen (Alibaba)
- Models
- 57
- Cheapest Input
- $0.030 /M tokens
🇺🇸 Google
- Models
- 51
- Cheapest Input
- $0.050 /M tokens
🇺🇸 Anthropic
- Models
- 36
- Cheapest Input
- $0.250 /M tokens
🇫🇷 Mistral AI
- Models
- 26
- Cheapest Input
- $0.019 /M tokens
🇨🇳 Z.ai (Zhipu)
- Models
- 19
- Cheapest Input
- $0.060 /M tokens
🇨🇳 DeepSeek
- Models
- 18
- Cheapest Input
- $0.030 /M tokens
🇺🇸 NVIDIA
- Models
- 17
- Cheapest Input
- $0.040 /M tokens
Last update: 2026-08-26 · Prices in USD, for reference only. Please verify with official sources.