The Frontier Index
The state of the AI frontier, in one place
40 frontier and open-weight models across 10 providers: the cheapest to run, the biggest context windows, the newest arrivals, and how far open weights now undercut proprietary APIs. Every figure derives from one hand-verified, versioned ledger, last updated 26 Sep 2026.
Cheapest input
GPT-6 Luna
$0.10 per 1M in · OpenAI · tied with one other
Cheapest to run
Llama 4 Scout
$0.15 / 1M · 3:1 blend
Longest context
GPT-6 Astra
1.05M tokens · OpenAI · tied with seven others
Newest arrival
Claude Opus 5.5
released 22 Sep 2026 · tied with two others
Open weights
11 of 40
self-hostable models tracked
Providers
10
40 models, one verified ledger
The gap has flipped
At the ~1M-token frontier, the cheapest proprietary model, GPT-6 Luna, runs about 1.2× cheaper than the cheapest open-weights one, GLM-5.3-Flash.
$0.20 vs $0.24 per 1M on a blended $/1M at a 3:1 input:output mix. Across all 40 models, blended list price spans $0.15 to $67.50 per 1M, a 450× range. The trade-off: open weights mean you run the infrastructure, and the top proprietary models still lead on some frontier reasoning.
Figures are base list rates for prompts up to ~200K tokens; some models add a long-context surcharge above that, and open-weights API rates are representative third-party-hosted prices, and self-hosting is free.
Pick by requirement
Set your real constraints (context window, open weights, multimodal input) and see the cheapest model in the ledger that clears them.
Cheapest that qualifies(40 models match)
- Llama 4 ScoutcheapestMeta320K$0.15/1M($0.10 in · $0.30 out)
- GPT-6 LunaOpenAI1.05M$0.20/1M($0.10 in · $0.50 out)
- Qwen 3.8 FlashAlibaba1M$0.23/1M($0.15 in · $0.47 out)
Ranked by blended $/1M at a 3:1 input:output mix. Prices are base list rates (up to ~200K tokens); open-weights API figures are representative hosted rates (self-hosting is free).
Every model, cheapest first
All 40 tracked models, ranked by blended $/1M at a 3:1 input:output mix: the blend weights input over output because assistant and agent workloads re-send a lot of context. Each row shows the date it was last verified and links straight to the provider's official pricing page.
Swipe for prices, context + source →
| # | Model | Blended $/1M | In · Out | Context | Verified | Source |
|---|---|---|---|---|---|---|
| 1 | Llama 4 ScoutMeta · open | $0.15* | $0.10 · $0.30 | 320K | 15 Sep 2026 | hosted ↗ |
| 2 | GPT-6 LunaOpenAI | $0.20* | $0.10 · $0.50 | 1.05M | 23 Sep 2026 | official ↗ |
| 3 | Qwen 3.8 FlashAlibaba | $0.23* | $0.15 · $0.47 | 1M | 22 Sep 2026 | official ↗ |
| 4 | Qwen 3.8 Omni FlashAlibaba | $0.23* | $0.15 · $0.47 | 1M | 23 Sep 2026 | official ↗ |
| 5 | GLM-5.3-FlashZhipu / Z.ai · open | $0.24* | $0.15 · $0.50 | 1M | 19 Sep 2026 | official ↗ |
| 6 | Mistral Small 4Mistral AI · open | $0.26* | $0.15 · $0.60 | 256K | 22 Sep 2026 | official ↗ |
| 7 | Llama 4 MaverickMeta · open | $0.35* | $0.20 · $0.80 | 1M | 15 Sep 2026 | hosted ↗ |
| 8 | GPT-5.6 LunaOpenAI | $0.45* | $0.20 · $1.20 | 1.05M | 3 Sep 2026 | official ↗ |
| 9 | GPT-5.4 nanoOpenAI | $0.46 | $0.20 · $1.25 | 400K | 3 Sep 2026 | official ↗ |
| 10 | DeepSeek V4.1 FlashDeepSeek · open | $0.52* | $0.30 · $1.20 | 1M | 10 Sep 2026 | official ↗ |
| 11 | GLM-5.3-FlashXZhipu / Z.ai | $0.59* | $0.37 · $1.25 | 1M | 19 Sep 2026 | official ↗ |
| 12 | Mistral Large 3Mistral AI · open | $0.75* | $0.50 · $1.50 | 256K | 3 Sep 2026 | official ↗ |
| 13 | Gemini 3.8 FlashGoogle | $1.50* | $0.75 · $3.75 | 1M | 22 Sep 2026 | official ↗ |
| 14 | GLM-5Zhipu / Z.ai · open | $1.55* | $1 · $3.20 | 200K | 3 Sep 2026 | official ↗ |
| 15 | Grok 4.3xAI | $1.56* | $1.25 · $2.50 | 1M | 3 Sep 2026 | official ↗ |
| 16 | GPT-5.4 miniOpenAI | $1.69 | $0.75 · $4.50 | 400K | 3 Sep 2026 | official ↗ |
| 17 | DeepSeek V4DeepSeek · open | $1.98* | $1.32 · $3.96 | 1M | 10 Sep 2026 | official ↗ |
| 18 | Claude Haiku 4.5Anthropic | $2 | $1 · $5 | 200K | 26 Sep 2026 | official ↗ |
| 19 | GLM-5.3Zhipu / Z.ai · open | $2.15* | $1.40 · $4.40 | 1M | 19 Sep 2026 | official ↗ |
| 20 | Grok 4.7xAI | $3* | $2 · $6 | 500K | 21 Sep 2026 | official ↗ |
| 21 | Grok 4.6xAI | $3* | $2 · $6 | 500K | 21 Sep 2026 | official ↗ |
| 22 | Qwen 3.8 MaxAlibaba | $3* | $2 · $6 | 1M | 3 Sep 2026 | official ↗ |
| 23 | Mistral Medium 3.5Mistral AI · open | $3* | $1.50 · $7.50 | 256K | 3 Sep 2026 | official ↗ |
| 24 | Gemini 3.5 FlashGoogle | $3.38* | $1.50 · $9 | 1M | 3 Sep 2026 | official ↗ |
| 25 | Qwen 3.7 MaxAlibaba | $3.75* | $2.50 · $7.50 | 1M | 3 Sep 2026 | official ↗ |
| 26 | Claude Sonnet 5Anthropic | $4* | $2 · $10 | 1M | 26 Sep 2026 | official ↗ |
| 27 | GPT-6 SolOpenAI | $4* | $2 · $10 | 1.05M | 23 Sep 2026 | official ↗ |
| 28 | GPT-5.6 TerraOpenAI | $4.50* | $2 · $12 | 1.05M | 3 Sep 2026 | official ↗ |
| 29 | Gemini 3.1 ProGoogle | $4.50* | $2 · $12 | 1M | 3 Sep 2026 | official ↗ |
| 30 | Claude Sonnet 4.6Anthropic | $6* | $3 · $15 | 1M | 25 Sep 2026 | official ↗ |
| 31 | Kimi K3Moonshot AI · open | $6* | $3 · $15 | 1M | 3 Sep 2026 | official ↗ |
| 32 | Claude Opus 5.5Anthropic | $8* | $4 · $20 | 1M | 22 Sep 2026 | official ↗ |
| 33 | GPT-5.6 SolOpenAI | $8* | $4 · $20 | 1.05M | 3 Sep 2026 | official ↗ |
| 34 | Claude Opus 5Anthropic | $10* | $5 · $25 | 1M | 3 Sep 2026 | official ↗ |
| 35 | Claude Opus 4.8Anthropic | $10 | $5 · $25 | 1M | 25 Sep 2026 | official ↗ |
| 36 | GPT-5.5OpenAI | $11.25* | $5 · $30 | 1.05M | 5 Sep 2026 | official ↗ |
| 37 | Claude Fable 5.1Anthropic | $20* | $10 · $50 | 1M | 26 Sep 2026 | official ↗ |
| 38 | Claude Fable 5Anthropic | $20 | $10 · $50 | 1M | 25 Sep 2026 | official ↗ |
| 39 | GPT-6 AstraOpenAI | $20* | $10 · $50 | 1.05M | 5 Sep 2026 | official ↗ |
| 40 | GPT-5.5 ProOpenAI | $67.50* | $30 · $180 | 1.05M | 22 Sep 2026 | official ↗ |
Blended figures are the base list rate for prompts up to ~200K tokens. * = a long-context surcharge applies above that (hover for the tier). Open-weights API rates are representative third-party-hosted prices; self-hosting is free. The verified date is when each row was last checked against its linked source.
Frequently asked
What's the cheapest AI model right now?
By raw input price, two models tie for lowest at $0.10 per 1M input tokens: GPT-6 Luna and Llama 4 Scout. But real workloads pay for output too: the cheapest to actually run, on a blended $/1M at a 3:1 input:output mix, is Llama 4 Scout at $0.15 per 1M. Open-weights models can also be self-hosted for the cost of your own hardware.
Which model has the largest context window?
Eight models share the top spot, all at 1.05M tokens: GPT-6 Astra, GPT-6 Sol, GPT-6 Luna, GPT-5.5 Pro, GPT-5.5, GPT-5.6 Sol, GPT-5.6 Terra and GPT-5.6 Luna. The 1M class is wider still: 30 of the 40 models tracked have a ~1,000K-token window, so very-long-context work is no longer a single-vendor feature.
Are open-weights models really cheaper than proprietary ones?
Not at the ~1M-token frontier any more — there the cheapest proprietary tier now undercuts the cheapest open one. At the ~1M-token frontier the cheapest proprietary model, GPT-6 Luna, runs about 1.2× cheaper than the cheapest open-weights one, GLM-5.3-Flash, on a blended $/1M at a 3:1 input:output mix ($0.20 vs $0.24 per 1M). Open weights still self-host for hardware cost alone, which no per-token rate can match, and the trade-off is that you run the infrastructure. The top proprietary models also still lead on some frontier reasoning tasks.
How current is this, and how do you verify it?
Every figure is hand-verified against published pricing and cross-checked for contradictions. Each row carries its own verification date and a link to the page the price is printed on (both shown in the full table below) — usually the provider's own pricing page, marked "official". A few models are open weights their maker gives away and sells no API for, so no first-party rate exists to cite; those rows show a named host's list price instead and are marked "hosted". The catalog was last updated 26 Sep 2026, and the last full cross-provider sweep was 3 Sep 2026; newer arrivals are verified on their own dates. It's a single, versioned source of truth (version 2026-09-26) that every BitByteCore tool reads, so a price can't drift between pages. Model pricing moves fast, so confirm with the provider before you rely on a figure.
Can I use or cite this data?
Yes. The full catalog is published as machine-readable JSON at https://bitbytecore.com/data/frontier.json and as schema.org Dataset markup on this page. Quote or cite it with clear attribution and a link back.
How this stays honest
One versioned catalog (v2026-09-26, last updated 26 Sep 2026) is the single source every BitByteCore model tool reads, so a price can’t drift between pages. It’s published openly for anyone to cite.
Put the numbers to work
- CalculatorAI API Cost CalculatorPlug in your own usage → monthly spend across every model, ranked cheapest-first.Open
- CounterToken CounterPaste text → exact token count, context-fit, and the cost to send it. Runs in your browser.Open
- CalculatorContext Window CalculatorWill your document fit? See which models can hold it in one window, and the cost.Open
- PickerWhich Model Should I Use?Answer a few questions → a top pick plus two alternatives, with the reasoning.Open