The Frontier Index
The state of the AI frontier, in one place
19 frontier and open-weight models across 10 providers — the cheapest to run, the biggest context windows, the newest arrivals, and how far open weights now undercut proprietary APIs. Every figure derives from one hand-verified, versioned ledger, last updated 20 Jul 2026.
Cheapest input
GPT-5.4 nano
$0.20 per 1M in · OpenAI
Cheapest to run
Llama 4 Maverick
$0.38 / 1M · 3:1 blend
Longest context
GPT-5.5
1.05M tokens · OpenAI
Newest arrival
GPT-5.6 Sol
released 9 Jul 2026
Open weights
5 of 19
self-hostable models tracked
Providers
10
19 models, one verified ledger
The number that matters
At the ~1M-token frontier, the cheapest open-weights model — Llama 4 Maverick — runs about 4.1× cheaper than the cheapest proprietary one, Grok 4.3.
$0.38 vs $1.56 per 1M on a blended $/1M at a 3:1 input:output mix. Across all 19 models, blended list price spans $0.38 to $20 per 1M — a 53× range. The trade-off: open weights mean you run the infrastructure, and the top proprietary models still lead on some frontier reasoning.
Figures are base list rates for prompts up to ~200K tokens; some models add a long-context surcharge above that, and open-weights API rates are representative third-party-hosted prices — self-hosting is free.
Pick by requirement
Set your real constraints — context window, open weights, multimodal input — and see the cheapest model in the ledger that clears them.
Cheapest that qualifies(19 models match)
- Llama 4 MaverickcheapestMeta1M$0.38/1M($0.22 in · $0.85 out)
- GPT-5.4 nanoOpenAI400K$0.46/1M($0.20 in · $1.25 out)
- DeepSeek V4DeepSeek1M$0.54/1M($0.43 in · $0.87 out)
Ranked by blended $/1M at a 3:1 input:output mix. Prices are base list rates (up to ~200K tokens); open-weights API figures are representative hosted rates (self-hosting is free).
Every model, cheapest first
Filter & sort all 19 →All 19 tracked models, ranked by blended $/1M at a 3:1 input:output mix— the blend weights input over output because assistant and agent workloads re-send a lot of context. Each row shows the date it was last verified and links straight to the provider's official pricing page.
Swipe for prices, context + source →
| # | Model | Blended $/1M | In · Out | Context | Verified | Source |
|---|---|---|---|---|---|---|
| 1 | Llama 4 MaverickMeta · open | $0.38* | $0.22 · $0.85 | 1M | 20 Jul 2026 | official ↗ |
| 2 | GPT-5.4 nanoOpenAI | $0.46 | $0.20 · $1.25 | 400K | 20 Jul 2026 | official ↗ |
| 3 | DeepSeek V4DeepSeek · open | $0.54* | $0.43 · $0.87 | 1M | 20 Jul 2026 | official ↗ |
| 4 | Mistral Large 3Mistral AI · open | $0.75 | $0.50 · $1.50 | 256K | 20 Jul 2026 | official ↗ |
| 5 | Kimi K2 ThinkingMoonshot AI · open | $1.07 | $0.60 · $2.50 | 256K | 21 Jun 2026 | official ↗ |
| 6 | GLM-5Zhipu / Z.ai · open | $1.55* | $1 · $3.20 | 200K | 20 Jul 2026 | official ↗ |
| 7 | Grok 4.3xAI | $1.56* | $1.25 · $2.50 | 1M | 20 Jul 2026 | official ↗ |
| 8 | GPT-5.4 miniOpenAI | $1.69 | $0.75 · $4.50 | 400K | 20 Jul 2026 | official ↗ |
| 9 | Claude Haiku 4.5Anthropic | $2 | $1 · $5 | 200K | 20 Jul 2026 | official ↗ |
| 10 | GPT-5.6 LunaOpenAI | $2.25* | $1 · $6 | 1.05M | 10 Jul 2026 | official ↗ |
| 11 | Gemini 3.5 FlashGoogle | $3.38 | $1.50 · $9 | 1M | 20 Jul 2026 | official ↗ |
| 12 | Qwen 3.7 MaxAlibaba | $3.75* | $2.50 · $7.50 | 1M | 20 Jul 2026 | official ↗ |
| 13 | Gemini 3.1 ProGoogle | $4.50* | $2 · $12 | 1M | 20 Jul 2026 | official ↗ |
| 14 | GPT-5.6 TerraOpenAI | $5.63* | $2.50 · $15 | 1.05M | 10 Jul 2026 | official ↗ |
| 15 | Claude Sonnet 4.6Anthropic | $6* | $3 · $15 | 1M | 20 Jul 2026 | official ↗ |
| 16 | Claude Opus 4.8Anthropic | $10 | $5 · $25 | 1M | 20 Jul 2026 | official ↗ |
| 17 | GPT-5.5OpenAI | $11.25* | $5 · $30 | 1.05M | 20 Jul 2026 | official ↗ |
| 18 | GPT-5.6 SolOpenAI | $11.25* | $5 · $30 | 1.05M | 10 Jul 2026 | official ↗ |
| 19 | Claude Fable 5Anthropic | $20 | $10 · $50 | 1M | 20 Jul 2026 | official ↗ |
Blended figures are the base list rate for prompts up to ~200K tokens. * = a long-context surcharge applies above that (hover for the tier). Open-weights API rates are representative third-party-hosted prices; self-hosting is free. The verified date is when each row was last checked against its linked source.
Frequently asked
What's the cheapest AI model right now?
By raw input price, GPT-5.4 nano (OpenAI) is lowest at $0.20 per 1M input tokens. But real workloads pay for output too — the cheapest to actually run, on a blended $/1M at a 3:1 input:output mix, is Llama 4 Maverick at $0.38 per 1M. Open-weights models can also be self-hosted for the cost of your own hardware.
Which model has the largest context window?
GPT-5.5 leads at 1.05M tokens. It's not alone at the top: 13 of the 19 models tracked have a 1M-class (~1,000K-token) window, so very-long-context work is no longer a single-vendor feature.
Are open-weights models really cheaper than proprietary ones?
At the ~1M-token frontier, yes: the cheapest open-weights model, Llama 4 Maverick, runs about 4.1× cheaper than the cheapest proprietary one, Grok 4.3, on a blended $/1M at a 3:1 input:output mix ($0.38 vs $1.56 per 1M). Open weights also self-host for hardware cost alone — the trade-off is that you run the infrastructure and the top proprietary models still lead on some frontier reasoning tasks.
How current is this, and how do you verify it?
Every figure is hand-verified against each provider's own published pricing and cross-checked for contradictions. Each row carries its own verification date and a link to the provider's official pricing page (both shown in the full table below). The catalog was last updated 20 Jul 2026, and the last full cross-provider sweep was 20 Jul 2026; newer arrivals are verified on their own dates. It's a single, versioned source of truth (version 2026-07-20) that every BitByteCore tool reads, so a price can't drift between pages. Model pricing moves fast, so confirm with the provider before you rely on a figure.
Can I use or cite this data?
Yes. The full catalog is published as machine-readable JSON at https://bitbytecore.com/data/frontier.json and as schema.org Dataset markup on this page. Quote or cite it with clear attribution and a link back.
How this stays honest
One versioned catalog (v2026-07-20, last updated 20 Jul 2026) is the single source every BitByteCore model tool reads — so a price can't drift between pages. It's published openly for anyone to cite.
Put the numbers to work
- CalculatorAI API Cost CalculatorPlug in your own usage → monthly spend across every model, ranked cheapest-first.Open
- CounterToken CounterPaste text → exact token count, context-fit, and the cost to send it. Runs in your browser.Open
- CalculatorContext Window CalculatorWill your document fit? See which models can hold it in one window, and the cost.Open
- PickerWhich Model Should I Use?Answer a few questions → a top pick plus two alternatives, with the reasoning.Open