Skip to content

The Frontier Index

The state of the AI frontier, in one place

19 frontier and open-weight models across 10 providers — the cheapest to run, the biggest context windows, the newest arrivals, and how far open weights now undercut proprietary APIs. Every figure derives from one hand-verified, versioned ledger, last updated 20 Jul 2026.

The number that matters

At the ~1M-token frontier, the cheapest open-weights model — Llama 4 Maverick — runs about 4.1× cheaper than the cheapest proprietary one, Grok 4.3.

$0.38 vs $1.56 per 1M on a blended $/1M at a 3:1 input:output mix. Across all 19 models, blended list price spans $0.38 to $20 per 1M — a 53× range. The trade-off: open weights mean you run the infrastructure, and the top proprietary models still lead on some frontier reasoning.

Figures are base list rates for prompts up to ~200K tokens; some models add a long-context surcharge above that, and open-weights API rates are representative third-party-hosted prices — self-hosting is free.

Pick by requirement

Set your real constraints — context window, open weights, multimodal input — and see the cheapest model in the ledger that clears them.

Context window
Constraints

Cheapest that qualifies(19 models match)

  • Llama 4 MaverickcheapestMeta1M$0.38/1M
  • GPT-5.4 nanoOpenAI400K$0.46/1M
  • DeepSeek V4DeepSeek1M$0.54/1M

Ranked by blended $/1M at a 3:1 input:output mix. Prices are base list rates (up to ~200K tokens); open-weights API figures are representative hosted rates (self-hosting is free).

Every model, cheapest first

Filter & sort all 19

All 19 tracked models, ranked by blended $/1M at a 3:1 input:output mix— the blend weights input over output because assistant and agent workloads re-send a lot of context. Each row shows the date it was last verified and links straight to the provider's official pricing page.

Swipe for prices, context + source →

#ModelBlended $/1MIn · OutContextVerifiedSource
1Llama 4 MaverickMeta · open$0.38*$0.22 · $0.851M20 Jul 2026official ↗
2GPT-5.4 nanoOpenAI$0.46$0.20 · $1.25400K20 Jul 2026official ↗
3DeepSeek V4DeepSeek · open$0.54*$0.43 · $0.871M20 Jul 2026official ↗
4Mistral Large 3Mistral AI · open$0.75$0.50 · $1.50256K20 Jul 2026official ↗
5Kimi K2 ThinkingMoonshot AI · open$1.07$0.60 · $2.50256K21 Jun 2026official ↗
6GLM-5Zhipu / Z.ai · open$1.55*$1 · $3.20200K20 Jul 2026official ↗
7Grok 4.3xAI$1.56*$1.25 · $2.501M20 Jul 2026official ↗
8GPT-5.4 miniOpenAI$1.69$0.75 · $4.50400K20 Jul 2026official ↗
9Claude Haiku 4.5Anthropic$2$1 · $5200K20 Jul 2026official ↗
10GPT-5.6 LunaOpenAI$2.25*$1 · $61.05M10 Jul 2026official ↗
11Gemini 3.5 FlashGoogle$3.38$1.50 · $91M20 Jul 2026official ↗
12Qwen 3.7 MaxAlibaba$3.75*$2.50 · $7.501M20 Jul 2026official ↗
13Gemini 3.1 ProGoogle$4.50*$2 · $121M20 Jul 2026official ↗
14GPT-5.6 TerraOpenAI$5.63*$2.50 · $151.05M10 Jul 2026official ↗
15Claude Sonnet 4.6Anthropic$6*$3 · $151M20 Jul 2026official ↗
16Claude Opus 4.8Anthropic$10$5 · $251M20 Jul 2026official ↗
17GPT-5.5OpenAI$11.25*$5 · $301.05M20 Jul 2026official ↗
18GPT-5.6 SolOpenAI$11.25*$5 · $301.05M10 Jul 2026official ↗
19Claude Fable 5Anthropic$20$10 · $501M20 Jul 2026official ↗

Blended figures are the base list rate for prompts up to ~200K tokens. * = a long-context surcharge applies above that (hover for the tier). Open-weights API rates are representative third-party-hosted prices; self-hosting is free. The verified date is when each row was last checked against its linked source.

Frequently asked

What's the cheapest AI model right now?

By raw input price, GPT-5.4 nano (OpenAI) is lowest at $0.20 per 1M input tokens. But real workloads pay for output too — the cheapest to actually run, on a blended $/1M at a 3:1 input:output mix, is Llama 4 Maverick at $0.38 per 1M. Open-weights models can also be self-hosted for the cost of your own hardware.

Which model has the largest context window?

GPT-5.5 leads at 1.05M tokens. It's not alone at the top: 13 of the 19 models tracked have a 1M-class (~1,000K-token) window, so very-long-context work is no longer a single-vendor feature.

Are open-weights models really cheaper than proprietary ones?

At the ~1M-token frontier, yes: the cheapest open-weights model, Llama 4 Maverick, runs about 4.1× cheaper than the cheapest proprietary one, Grok 4.3, on a blended $/1M at a 3:1 input:output mix ($0.38 vs $1.56 per 1M). Open weights also self-host for hardware cost alone — the trade-off is that you run the infrastructure and the top proprietary models still lead on some frontier reasoning tasks.

How current is this, and how do you verify it?

Every figure is hand-verified against each provider's own published pricing and cross-checked for contradictions. Each row carries its own verification date and a link to the provider's official pricing page (both shown in the full table below). The catalog was last updated 20 Jul 2026, and the last full cross-provider sweep was 20 Jul 2026; newer arrivals are verified on their own dates. It's a single, versioned source of truth (version 2026-07-20) that every BitByteCore tool reads, so a price can't drift between pages. Model pricing moves fast, so confirm with the provider before you rely on a figure.

Can I use or cite this data?

Yes. The full catalog is published as machine-readable JSON at https://bitbytecore.com/data/frontier.json and as schema.org Dataset markup on this page. Quote or cite it with clear attribution and a link back.

How this stays honest

One versioned catalog (v2026-07-20, last updated 20 Jul 2026) is the single source every BitByteCore model tool reads — so a price can't drift between pages. It's published openly for anyone to cite.

Put the numbers to work