Live data
Assembling the frontier
Ranking the latest models and provider endpoints.
Loading model dataLive data
Ranking the latest models and provider endpoints.
Loading model dataSide by side
Select up to 4 models to analyze quality, performance, pricing, and benchmarks.
Quick picks from top 10 models by Quality Index
Highest Quality
Nemotron 3.5 Lightning
23.6
Fastest Output
Nemotron 3.5 Lightning
529 tok/s
Best Value
Nemotron 3.5 Lightning
$0.09/M
Quality, Speed, Context, and Value at a glance
| Metric | Nemotron 3.5 Lightning |
|---|---|
| Creator | NVIDIA |
| Quality Index | 23.6👑 |
| Price per Million Tokens | $0.09💰 |
| Output Speed | 529 tok/s⚡ |
| Context Window | 262K📚 |
| Latency (TTFT) | 0.40s🚀 |
| Provider | Fireworks (BF16)4 more providers |
Estimated cost based on 1 million tokens per day usage
Nemotron 3.5 Lightning
$3
/month
Use coding, reasoning, math, and tool-use benchmarks to see where a model is actually strong instead of relying on a single overall score. A model that leads in quality may still be wrong for your workflow if your primary constraint is latency or cost.
The same model can be cheap on one host and expensive on another, or fast on one provider and unusable on the next. If the model looks promising, move to provider comparison before you commit.
Ranking library
Focused rankings for the decisions engineers actually make.