Find the Cheapest & Best-Fit LLM API

Real-time tracking of 63 providers, 560 models' pricing, capabilities & context window. Covering international and Chinese models, auto-updated daily at 04:00.

Tracked Models 560
Free Models 57
Chinese Provider Models 167

This site tracks public LLM API pricing and benchmarks with daily updates. Background, sources, and caveats live on the About page →

Best Value Models

Filters to models with AA Intelligence Index 40 or higher, then ranks by intelligence per dollar (AA score ÷ input price)

# Model Provider Input Benchmark
1 Upstage: Solar Pro 4 🇰🇷 Upstage $0.030 AA Index 42
2 DeepSeek: DeepSeek V4 Flash 0731 🇨🇳 DeepSeek $0.045 AA Index 52
3 Z.ai: GLM 5.3 Flash Multimodal 🇨🇳 Z.ai (Zhipu) $0.075 AA Index 57
4 DeepSeek: DeepSeek V4 Flash 0423 🇨🇳 DeepSeek $0.082 AA Index 42
5 Tencent: Hy3 🇨🇳 Tencent $0.083 AA Index 42

Strongest Models

Ranked by Artificial Analysis Intelligence Index, a cross-domain intelligence metric

# Model Provider Input AA Index
1 Claude Opus 5 🇺🇸 Anthropic $5.00 63
2 Anthropic: Claude Fable 5 🇺🇸 Anthropic $10.00 62
3 OpenAI: GPT-5.6 Sol 🇺🇸 OpenAI $2.00 61
4 SpaceXAI: Grok 4.6 🇺🇸 xAI $2.00 61
5 MoonshotAI: Kimi K3 🇨🇳 Moonshot AI $3.00 60

Cheapest Input Price

Cost in USD per million input tokens (excluding free models)

# Model Provider Input Output
1 inclusionAI: Ling-2.6-flash 🇨🇳 InclusionAI $0.010 $0.030
2 IBM: Granite 4.0 Micro 🇺🇸 IBM Granite $0.017 $0.112
3 Mistral: Mistral Nemo 🇫🇷 Mistral AI $0.019 $0.030
4 Ling-3.0-flash 🇨🇳 InclusionAI $0.021 $0.063
5 Nex AGI: Nex-N2-Mini 🇺🇸 Nexa AI $0.025 $0.100

Fastest Output

Ranked by Artificial Analysis measured output speed (tokens/sec)

# Model Provider t/s
1 Inception: Mercury 2 🇺🇸 Inception 928
2 Ling-3.0-flash 🇨🇳 InclusionAI 370
3 Google: Gemini 3.5 Flash Lite 🇺🇸 Google 368
4 Google: Gemini 2.5 Flash Lite 🇺🇸 Google 367
5 NVIDIA: Nemotron 3 Nano Omni (free) 🇺🇸 NVIDIA 326

Free Models (57)

Models with $0 input price. Some may still charge for output — click to view full pricing details.

Model Provider Output $/M Context
Arcee AI: Trinity Large Thinking (free) 🇺🇸 Arcee AI Free 262K
Baidu Qianfan: CoBuddy (free) 🇨🇳 Baidu Free 131K
Baidu: Qianfan-OCR-Fast (free) 🇨🇳 Baidu Free 66K
Cohere: North Mini Code (free) 🇨🇦 Cohere Free 256K
DeepSeek: DeepSeek V4 Flash (free) 🇨🇳 DeepSeek Free 1.05M

Top Providers

🇺🇸 OpenAI

Models
110
Cheapest Input
$0.025 /M tokens

🇨🇳 Qwen (Alibaba)

Models
60
Cheapest Input
$0.030 /M tokens

🇺🇸 Google

Models
52
Cheapest Input
$0.050 /M tokens

🇺🇸 Anthropic

Models
36
Cheapest Input
$0.250 /M tokens

🇫🇷 Mistral AI

Models
32
Cheapest Input
$0.019 /M tokens

🇨🇳 DeepSeek

Models
20
Cheapest Input
$0.030 /M tokens

🇨🇳 Z.ai (Zhipu)

Models
20
Cheapest Input
$0.060 /M tokens

🇺🇸 NVIDIA

Models
17
Cheapest Input
$0.040 /M tokens

Last update: 2026-08-29 · Prices in USD, for reference only. Please verify with official sources.