Skip to content

The Frontier Index

The state of the AI frontier, in one place

40 frontier and open-weight models across 10 providers: the cheapest to run, the biggest context windows, the newest arrivals, and how far open weights now undercut proprietary APIs. Every figure derives from one hand-verified, versioned ledger, last updated 26 Sep 2026.

The gap has flipped

At the ~1M-token frontier, the cheapest proprietary model, GPT-6 Luna, runs about 1.2× cheaper than the cheapest open-weights one, GLM-5.3-Flash.

$0.20 vs $0.24 per 1M on a blended $/1M at a 3:1 input:output mix. Across all 40 models, blended list price spans $0.15 to $67.50 per 1M, a 450× range. The trade-off: open weights mean you run the infrastructure, and the top proprietary models still lead on some frontier reasoning.

Figures are base list rates for prompts up to ~200K tokens; some models add a long-context surcharge above that, and open-weights API rates are representative third-party-hosted prices, and self-hosting is free.

Pick by requirement

Set your real constraints (context window, open weights, multimodal input) and see the cheapest model in the ledger that clears them.

Context window
Constraints

Cheapest that qualifies(40 models match)

  • Llama 4 ScoutcheapestMeta320K$0.15/1M
  • GPT-6 LunaOpenAI1.05M$0.20/1M
  • Qwen 3.8 FlashAlibaba1M$0.23/1M

Ranked by blended $/1M at a 3:1 input:output mix. Prices are base list rates (up to ~200K tokens); open-weights API figures are representative hosted rates (self-hosting is free).

Every model, cheapest first

All 40 tracked models, ranked by blended $/1M at a 3:1 input:output mix: the blend weights input over output because assistant and agent workloads re-send a lot of context. Each row shows the date it was last verified and links straight to the provider's official pricing page.

Filter & sort all 40

Swipe for prices, context + source →

#ModelBlended $/1MIn · OutContextVerifiedSource
1Llama 4 ScoutMeta · open$0.15*$0.10 · $0.30320K15 Sep 2026hosted ↗
2GPT-6 LunaOpenAI$0.20*$0.10 · $0.501.05M23 Sep 2026official ↗
3Qwen 3.8 FlashAlibaba$0.23*$0.15 · $0.471M22 Sep 2026official ↗
4Qwen 3.8 Omni FlashAlibaba$0.23*$0.15 · $0.471M23 Sep 2026official ↗
5GLM-5.3-FlashZhipu / Z.ai · open$0.24*$0.15 · $0.501M19 Sep 2026official ↗
6Mistral Small 4Mistral AI · open$0.26*$0.15 · $0.60256K22 Sep 2026official ↗
7Llama 4 MaverickMeta · open$0.35*$0.20 · $0.801M15 Sep 2026hosted ↗
8GPT-5.6 LunaOpenAI$0.45*$0.20 · $1.201.05M3 Sep 2026official ↗
9GPT-5.4 nanoOpenAI$0.46$0.20 · $1.25400K3 Sep 2026official ↗
10DeepSeek V4.1 FlashDeepSeek · open$0.52*$0.30 · $1.201M10 Sep 2026official ↗
11GLM-5.3-FlashXZhipu / Z.ai$0.59*$0.37 · $1.251M19 Sep 2026official ↗
12Mistral Large 3Mistral AI · open$0.75*$0.50 · $1.50256K3 Sep 2026official ↗
13Gemini 3.8 FlashGoogle$1.50*$0.75 · $3.751M22 Sep 2026official ↗
14GLM-5Zhipu / Z.ai · open$1.55*$1 · $3.20200K3 Sep 2026official ↗
15Grok 4.3xAI$1.56*$1.25 · $2.501M3 Sep 2026official ↗
16GPT-5.4 miniOpenAI$1.69$0.75 · $4.50400K3 Sep 2026official ↗
17DeepSeek V4DeepSeek · open$1.98*$1.32 · $3.961M10 Sep 2026official ↗
18Claude Haiku 4.5Anthropic$2$1 · $5200K26 Sep 2026official ↗
19GLM-5.3Zhipu / Z.ai · open$2.15*$1.40 · $4.401M19 Sep 2026official ↗
20Grok 4.7xAI$3*$2 · $6500K21 Sep 2026official ↗
21Grok 4.6xAI$3*$2 · $6500K21 Sep 2026official ↗
22Qwen 3.8 MaxAlibaba$3*$2 · $61M3 Sep 2026official ↗
23Mistral Medium 3.5Mistral AI · open$3*$1.50 · $7.50256K3 Sep 2026official ↗
24Gemini 3.5 FlashGoogle$3.38*$1.50 · $91M3 Sep 2026official ↗
25Qwen 3.7 MaxAlibaba$3.75*$2.50 · $7.501M3 Sep 2026official ↗
26Claude Sonnet 5Anthropic$4*$2 · $101M26 Sep 2026official ↗
27GPT-6 SolOpenAI$4*$2 · $101.05M23 Sep 2026official ↗
28GPT-5.6 TerraOpenAI$4.50*$2 · $121.05M3 Sep 2026official ↗
29Gemini 3.1 ProGoogle$4.50*$2 · $121M3 Sep 2026official ↗
30Claude Sonnet 4.6Anthropic$6*$3 · $151M25 Sep 2026official ↗
31Kimi K3Moonshot AI · open$6*$3 · $151M3 Sep 2026official ↗
32Claude Opus 5.5Anthropic$8*$4 · $201M22 Sep 2026official ↗
33GPT-5.6 SolOpenAI$8*$4 · $201.05M3 Sep 2026official ↗
34Claude Opus 5Anthropic$10*$5 · $251M3 Sep 2026official ↗
35Claude Opus 4.8Anthropic$10$5 · $251M25 Sep 2026official ↗
36GPT-5.5OpenAI$11.25*$5 · $301.05M5 Sep 2026official ↗
37Claude Fable 5.1Anthropic$20*$10 · $501M26 Sep 2026official ↗
38Claude Fable 5Anthropic$20$10 · $501M25 Sep 2026official ↗
39GPT-6 AstraOpenAI$20*$10 · $501.05M5 Sep 2026official ↗
40GPT-5.5 ProOpenAI$67.50*$30 · $1801.05M22 Sep 2026official ↗

Blended figures are the base list rate for prompts up to ~200K tokens. * = a long-context surcharge applies above that (hover for the tier). Open-weights API rates are representative third-party-hosted prices; self-hosting is free. The verified date is when each row was last checked against its linked source.

Frequently asked

What's the cheapest AI model right now?

By raw input price, two models tie for lowest at $0.10 per 1M input tokens: GPT-6 Luna and Llama 4 Scout. But real workloads pay for output too: the cheapest to actually run, on a blended $/1M at a 3:1 input:output mix, is Llama 4 Scout at $0.15 per 1M. Open-weights models can also be self-hosted for the cost of your own hardware.

Which model has the largest context window?

Eight models share the top spot, all at 1.05M tokens: GPT-6 Astra, GPT-6 Sol, GPT-6 Luna, GPT-5.5 Pro, GPT-5.5, GPT-5.6 Sol, GPT-5.6 Terra and GPT-5.6 Luna. The 1M class is wider still: 30 of the 40 models tracked have a ~1,000K-token window, so very-long-context work is no longer a single-vendor feature.

Are open-weights models really cheaper than proprietary ones?

Not at the ~1M-token frontier any more — there the cheapest proprietary tier now undercuts the cheapest open one. At the ~1M-token frontier the cheapest proprietary model, GPT-6 Luna, runs about 1.2× cheaper than the cheapest open-weights one, GLM-5.3-Flash, on a blended $/1M at a 3:1 input:output mix ($0.20 vs $0.24 per 1M). Open weights still self-host for hardware cost alone, which no per-token rate can match, and the trade-off is that you run the infrastructure. The top proprietary models also still lead on some frontier reasoning tasks.

How current is this, and how do you verify it?

Every figure is hand-verified against published pricing and cross-checked for contradictions. Each row carries its own verification date and a link to the page the price is printed on (both shown in the full table below) — usually the provider's own pricing page, marked "official". A few models are open weights their maker gives away and sells no API for, so no first-party rate exists to cite; those rows show a named host's list price instead and are marked "hosted". The catalog was last updated 26 Sep 2026, and the last full cross-provider sweep was 3 Sep 2026; newer arrivals are verified on their own dates. It's a single, versioned source of truth (version 2026-09-26) that every BitByteCore tool reads, so a price can't drift between pages. Model pricing moves fast, so confirm with the provider before you rely on a figure.

Can I use or cite this data?

Yes. The full catalog is published as machine-readable JSON at https://bitbytecore.com/data/frontier.json and as schema.org Dataset markup on this page. Quote or cite it with clear attribution and a link back.

How this stays honest

One versioned catalog (v2026-09-26, last updated 26 Sep 2026) is the single source every BitByteCore model tool reads, so a price can’t drift between pages. It’s published openly for anyone to cite.

Put the numbers to work