Not just the cheapestโthe best value. We rank 157 AI models by quality-per-dollar using live pricing data from 79 providers.
Historical snapshot
This page is a dated monthly snapshot. For the live version that is better aligned to current rankings and search intent, use Best AI Models (Live) or jump to Best Open Source LLM.
Value Score = Quality Index รท Price per Million Tokens
Value Score
460.4
DeepSeek via DeepInfra
Value Score
370.8
OpenAI via CoreWeave
Value Score
340.0
Alibaba via DeepInfra (FP8)
| Rank | Model | Value | Quality | Price/M | Speed | Provider |
|---|---|---|---|---|---|---|
| 1 | DeepSeek V4 Flash 0731 (Reasoning, Max Effort) DeepSeek | 460.4 | 51.8 | $0.113 | 59 tok/s | DeepInfra |
| 2 | gpt-oss-120b (high) OpenAI | 370.8 | 24.1 | $0.065 | 36 tok/s | CoreWeave |
| 3 | Qwen3.5 4B (Reasoning) Alibaba | 340.0 | 20.4 | $0.060 | 32 tok/s | DeepInfra (FP8) |
| 4 | Ling 3.0 Flash InclusionAI | 339.8 | 37.8 | $0.111 | 427 tok/s | DeepInfra |
| 5 | Gemma 4 E4B (Reasoning) | 310.0 | 12.4 | $0.040 | 89 tok/s | DeepInfra |
| 6 | HyperNova 60B 2605 Multiverse Computing | 281.5 | 18.3 | $0.065 | 384 tok/s | CompactifAI |
| 7 | gpt-oss-20b (high) OpenAI | 276.4 | 15.2 | $0.055 | 89 tok/s | CoreWeave |
| 8 | Qwen3.5 4B (Non-reasoning) Alibaba | 268.3 | 16.1 | $0.060 | 27 tok/s | DeepInfra FP8 |
| 9 | gpt-oss-20b (low) OpenAI | 261.8 | 14.4 | $0.055 | 87 tok/s | CoreWeave |
| 10 | gpt-oss-120b (low) OpenAI | 229.2 | 14.9 | $0.065 | 30 tok/s | CoreWeave |
Top models for high-volume, cost-sensitive workloads.
Self-hostable models with best API value when you can't self-host.
DeepSeek V4 Flash 0731 (Reasoning, Max Effort)
DeepSeek via DeepInfra
$0.113
Value: 460.4
gpt-oss-120b (high)
OpenAI via CoreWeave
$0.065
Value: 370.8
Qwen3.5 4B (Reasoning)
Alibaba via DeepInfra (FP8)
$0.060
Value: 340.0
Ling 3.0 Flash
InclusionAI via DeepInfra
$0.111
Value: 339.8
Gemma 4 E4B (Reasoning)
Google via DeepInfra
$0.040
Value: 310.0
Use our interactive explorer to compare pricing across all 79 providers. Filter by quality, speed, and price to find your perfect model.
As of January 2026, DeepSeek V4 Flash 0731 (Reasoning, Max Effort) offers the best value under $1/M at $0.113/M with a quality index of 51.8. For absolute lowest cost, open source models like DeepSeek and Qwen can be self-hosted for near-zero marginal cost after infrastructure.
DeepSeek V4 Flash 0731 (Reasoning, Max Effort) currently offers the best quality-per-dollar with a value score of 460.4(Quality Index 51.8 at $0.113/M). This means you get more "intelligence" per dollar spent than any other model.
For most production workloads, no. Models like DeepSeek V3.2 and Gemini Flash deliver 85-95% of GPT-5's quality at 1/10th to 1/20th the cost. Use GPT-5 for: (1) complex multi-step reasoning, (2) tasks where error cost is high, (3) when you need specific OpenAI features. Use budget models for: high-volume chat, content generation, and routine tasks.
Use API if: you have <1M tokens/day, need instant scaling, or lack GPU infrastructure.Self-host if: you have >10M tokens/day (break-even point), need data privacy, or want to fine-tune. At current GPU prices, self-hosting DeepSeek V3 becomes cheaper than API around 5-10M tokens/day depending on your setup.
Self-hostable models
๐ปLiveCodeBench leaders
๐คTool use & agents
๐Expert picks
Data sources: Pricing from Artificial Analysis (live API data). Quality Index from AA Intelligence Index. Updated daily via automated pipeline.See methodology โ