Find the Cheapest & Best-Fit LLM API
Real-time tracking of 66 providers, 620 models' pricing, capabilities & context window. Covering international and Chinese models, auto-updated daily at 04:00.
This site tracks public LLM API pricing and benchmarks with daily updates. Background, sources, and caveats live on the About page →
Best Value Models
Filters to models with AA Intelligence Index 40 or higher, then ranks by intelligence per dollar (AA score ÷ input price)
| # | Model | Provider | Input | Benchmark |
|---|---|---|---|---|
| 1 | Z.ai: GLM 5.3 Flash Multimodal | | $0.150 | AA Index 42 |
| 2 | Z.ai: GLM 5.3 | | $0.654 | AA Index 45 |
| 3 | Google: Gemini 3.8 Flash MultimodalVoice | | $0.750 | AA Index 41 |
| 4 | Meta: Muse Spark 1.3 MultimodalVoice | meta | $1.25 | AA Index 45 |
| 5 | OpenAI: GPT-5.6 Sol Multimodal | | $2.00 | AA Index 47 |
Strongest Models
Ranked by Artificial Analysis Intelligence Index, a cross-domain intelligence metric
| # | Model | Provider | Input | AA Index |
|---|---|---|---|---|
| 1 | Anthropic: Claude Fable 5.1 | | $10.00 | 53 |
| 2 | OpenAI: GPT-6 Astra | | $10.00 | 53 |
| 3 | Anthropic: Claude Opus 5 | | $5.00 | 51 |
| 4 | Anthropic: Claude Fable 5 | | $10.00 | 50 |
| 5 | OpenAI: GPT-5.6 Sol | | $2.00 | 47 |
Cheapest Input Price
Cost in USD per million input tokens (excluding free models)
| # | Model | Provider | Input | Output |
|---|---|---|---|---|
| 1 | inclusionAI: Ling-2.6-flash | | $0.010 | $0.030 |
| 2 | IBM: Granite 4.0 Micro | | $0.017 | $0.112 |
| 3 | OpenAI: gpt-oss-20b | | $0.018 | $0.090 |
| 4 | Mistral: Mistral Nemo | | $0.019 | $0.030 |
| 5 | inclusionAI: Ling 3.0 Flash | | $0.021 | $0.063 |
Fastest Output
Ranked by Artificial Analysis measured output speed (tokens/sec)
| # | Model | Provider | t/s |
|---|---|---|---|
| 1 | Inception: Mercury 2 | | 925 |
| 2 | Google: Gemini 2.5 Flash Lite | | 401 |
| 3 | Google: Gemini 3.5 Flash Lite | | 363 |
| 4 | Google: Gemini 3.7 Flash | | 322 |
| 5 | Arcee AI: Trinity Large Thinking | | 320 |
Free Models (64)
Models with $0 input price. Some may still charge for output — click to view full pricing details.
| Model | Provider | Output $/M | Context |
|---|---|---|---|
| Arcee AI: Trinity Large Thinking (free) | | Free | 262K |
| Baidu Qianfan: CoBuddy (free) | | Free | 131K |
| Baidu: Qianfan-OCR-Fast (free) | | Free | 66K |
| Cohere: North Mini Code (free) | | Free | 256K |
| DeepSeek: DeepSeek V4 Flash (free) | | Free | 1.05M |
Top Providers
🇺🇸 OpenAI
- Models
- 126
- Cheapest Input
- $0.018 /M tokens
🇨🇳 Qwen (Alibaba)
- Models
- 63
- Cheapest Input
- $0.030 /M tokens
🇺🇸 Google
- Models
- 54
- Cheapest Input
- $0.050 /M tokens
🇺🇸 Anthropic
- Models
- 40
- Cheapest Input
- $0.250 /M tokens
🇫🇷 Mistral AI
- Models
- 32
- Cheapest Input
- $0.019 /M tokens
🇨🇳 DeepSeek
- Models
- 26
- Cheapest Input
- $0.030 /M tokens
🇨🇳 Z.ai (Zhipu)
- Models
- 23
- Cheapest Input
- $0.060 /M tokens
🇺🇸 NVIDIA
- Models
- 18
- Cheapest Input
- $0.040 /M tokens
Last update: 2026-09-22 · Prices in USD, for reference only. Please verify with official sources.