Models
The labs keep shipping. This page is the current flagship cut: closed APIs from OpenAI, Anthropic, Google, xAI, and Meta, plus the open-weight models that actually sit next to them on a shared harness — Kimi, GLM, Qwen, DeepSeek, Gemma, Mistral, Llama. Quality is the Artificial Analysis Intelligence Index (mostly v4.1.1; Muse Spark 1.3, Gemini 3.8 Flash, and DeepSeek V4.1 Flash on v4.3), not a press-release table.
Prices are USD per million tokens on the first-party (or AA-measured) API. If a field is missing, it is a dash, not a guess. Llama 4 is still here because that is the downloadable Meta line; Muse Spark 1.3 is the Meta API that currently scores on AA. We do not treat every 7B fine-tune as a frontier model.
Price vs valueAPI tiers: list $/1M next to what you get.Plan tiersChat/product subscriptions: Free → Enterprise.ArenaMultiple Arena boards as charts — Text, Code, Vision, Search, Agent.LiveBenchCategory charts + full table — contamination-resistant.Pricing historyInput and output $/1M over time — step chart.Usage AdvisorReprice your API mix; CSV stays in-browser.What ifTools: past stock buys.
Higher is better. Navy is frontier / API-only. Accent is open weight. Effort settings are the ones AA scored, not always the lab default. Bars are scaled to this board's score range (not from zero).
- Claude Fable 5.166
- Claude Mythos 5.166
- Claude Opus 563
- Claude Fable 562
- GPT-6 Astra61
- GPT-5.6 Sol61
- Grok 4.661
- Kimi K360
- GLM-5.360
- Qwen3.8 2.4T A95B58
- GLM-5.3-Flash57
- GPT-5.6 Terra57
- Muse Spark 1.257
- Gemini 3.7 Flash56
- Claude Sonnet 555
- DeepSeek V4 Pro 081353
- DeepSeek V4 Flash 073152
- GPT-5.6 Luna52
- Muse Spark 1.348
- Gemini 3.1 Pro Preview48
- Grok 4.746
- Gemini 3.8 Flash41
- DeepSeek V4.1 Flash40
- Qwen3.5 397B A17B Int434
- Claude Haiku 4.530
- Gemma 4 31B30
- Mistral Medium 3.530
- Llama 4 Maverick14
Filter
- Claude Fable 5.1
Anthropic · Frontier closed
Anthropic's Sep 2026 coding and research flagship; tops the current AA index.
- Index
- 66
- Context
- 1M
- In / 1M
- $10
- Out / 1M
- $50
- License
- —
- Released
- 1 Sep 2026
Scored: adaptive reasoning, max effort, default fallback
Cache reads $0.25 per 1M tokens (0.025× input). Input/output list price unchanged from Fable 5.
- Claude Mythos 5.1
Anthropic · Frontier closed
Invite-only (not publicly available); same weights as Fable 5.1; more permissive cyber/life-sciences safeguards via Project Glasswing / trusted access.
- Index
- 66
- Context
- 1M
- In / 1M
- $10
- Out / 1M
- $50
- License
- —
- Released
- 1 Sep 2026
Scored: AA scored Fable 5.1 (identical weights); adaptive reasoning, max effort, default fallback
Not public API access — Project Glasswing / trusted access only. Cache reads $0.25 per 1M tokens. List price matches Fable 5.1 where invited.
- Claude Opus 5
Anthropic · Frontier closed
Anthropic's current flagship for long-horizon agents and deep reasoning.
- Index
- 63
- Context
- 1M
- In / 1M
- $5
- Out / 1M
- $25
- License
- —
- Released
- 24 Jul 2026
Scored: adaptive reasoning, max effort
- Claude Fable 5
Anthropic · Frontier closed
Highest-priced Claude in the current lineup; reserved for the heaviest agent jobs.
- Index
- 62
- Context
- 1M
- In / 1M
- $10
- Out / 1M
- $50
- License
- —
- Released
- 9 Jun 2026
Scored: adaptive reasoning, max effort, Opus 4.8 fallback
- GPT-6 Astra
OpenAI · Frontier closed
OpenAI's GPT-6 flagship for hard end-to-end agent work; Critical-cyber tier with gated tools.
- Index
- 61
- Context
- 1.05M
- In / 1M
- $10
- Out / 1M
- $50
- License
- —
- Released
- 3 Sep 2026
Scored: max effort
Short-context list. Long context bills $20 / $75 per 1M. Cache hit $1.00 / $2.00 short/long. Rolling out via Trusted Access first.
- GPT-5.6 Sol
OpenAI · Frontier closed
OpenAI's GPT-5.6 flagship; the gpt-5.6 alias routes here.
- Index
- 61
- Context
- 1.05M
- In / 1M
- $4
- Out / 1M
- $20
- License
- —
- Released
- 9 Jul 2026
Scored: max effort
Promotional list price through at least 21 Nov 2026. Prompts over 272k input tokens bill 2× input and 1.5× output.
- Grok 4.6
xAI · Frontier closed
Prior xAI coding/agent flagship on a 500k window; prefer Grok 4.7 for the current cut.
- Index
- 61
- Context
- 500k
- In / 1M
- $2
- Out / 1M
- $6
- License
- —
- Released
- 12 Aug 2026
Scored: high effort
Standard band under 200k prompt tokens, per xAI docs.
- Kimi K3
Moonshot / Kimi · Open weight
Highest-ranked open-weight model on the current Intelligence Index (2.8T MoE, 104B active).
- Index
- 60
- Context
- 1.05M
- In / 1M
- $3
- Out / 1M
- $15
- License
- Kimi K3 License (commercial with restrictions)
- Released
- 16 Jul 2026
Scored: max effort
- GLM-5.3
Zhipu / Z.AI · Open weight
Zhipu's current open-weight frontier; ties Kimi K3 at 60 on the Index.
- Index
- 60
- Context
- 1M
- In / 1M
- $1.40
- Out / 1M
- $4.40
- License
- GLM-5 license
- Released
- 14 Aug 2026
Scored: max effort
- Qwen3.8 2.4T A95B
Alibaba / Qwen · Open weight
Alibaba's 2.4T MoE open-weight flagship (95B active, per the AA card).
- Index
- 58
- Context
- 984k
- In / 1M
- $2
- Out / 1M
- $6
- License
- Qwen3 license
- Released
- 12 Aug 2026
- GLM-5.3-Flash
Zhipu / Z.AI · Open weight
Z.ai’s multimodal Flash cut of GLM-5.3 — 320B MoE / 18B active; ties GPT-5.6 Terra and Muse Spark at 57 on far cheaper tokens.
- Index
- 57
- Context
- 1M
- In / 1M
- $0.15
- Out / 1M
- $0.50
- License
- MIT
- Released
- August 2026
Scored: max effort
Native text+image input. ~10× cheaper list than GLM-5.3 ($1.40/$4.40).
- GPT-5.6 Terra
OpenAI · Frontier closed
Mid GPT-5.6 tier: most of Sol's window at half the input price.
- Index
- 57
- Context
- 1.05M
- In / 1M
- $2
- Out / 1M
- $12
- License
- —
- Released
- 9 Jul 2026
Scored: max effort
- Muse Spark 1.2
Meta · Frontier closed
Previous Meta API quality cut; kept for Index v4.1.1 comparison. Prefer Muse Spark 1.3 for the current Meta scorer.
- Index
- 57
- Context
- 1.05M
- In / 1M
- $1.25
- Out / 1M
- $4.25
- License
- —
- Released
- 5 Aug 2026
Scored: xhigh
Prior Meta API scorer on Index v4.1.1; superseded by Muse Spark 1.3 for current Meta quality.
- Gemini 3.7 Flash
Google · Frontier closed
Google's current high-throughput Gemini; scores above 3.1 Pro Preview on this Index.
- Index
- 56
- Context
- 1M
- In / 1M
- $0.75
- Out / 1M
- $3.75
- License
- —
- Released
- 13 Aug 2026
Scored: high
- Claude Sonnet 5
Anthropic · Frontier closed
Workhorse Claude: 1M context at Sonnet prices.
- Index
- 55
- Context
- 1M
- In / 1M
- $2
- Out / 1M
- $10
- License
- —
- Released
- 30 Jun 2026
Scored: adaptive reasoning, max effort
- DeepSeek V4 Pro 0813
DeepSeek · Open weight
1.6T MoE, 49B active, MIT weights; the current DeepSeek quality cut.
- Index
- 53
- Context
- 1M
- In / 1M
- $1.32
- Out / 1M
- $3.96
- License
- MIT
- Released
- 13 Aug 2026
Scored: max effort
First-party API as measured by Artificial Analysis (may differ from off-peak list).
- DeepSeek V4 Flash 0731
DeepSeek · Open weight
Cheaper DeepSeek V4 cut that still clears 50 on the Index.
- Index
- 52
- Context
- 1M
- In / 1M
- $0.44
- Out / 1M
- $1.32
- License
- MIT
- Released
- 31 Jul 2026
Scored: max effort
First-party API as measured by Artificial Analysis.
- GPT-5.6 Luna
OpenAI · Frontier closed
Volume GPT-5.6 tier: same 1.05M window, priced for high-throughput work.
- Index
- 52
- Context
- 1.05M
- In / 1M
- $0.20
- Out / 1M
- $1.20
- License
- —
- Released
- 9 Jul 2026
Scored: max effort
- Muse Spark 1.3
Meta · Frontier closed
Meta's current API quality line (proprietary, not Llama weights); AA Index 48 on v4.3 at max reasoning.
- Index
- 48
- Context
- 1M
- In / 1M
- $1.25
- Out / 1M
- $4.25
- License
- —
- Released
- 2 Sep 2026
Scored: max
Standard muse-spark-1.3 route (not used to improve products). Cached input $0.15 per 1M. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board.
- Gemini 3.1 Pro Preview
Google · Frontier closed
Google's Pro-class preview; Flash 3.7 now outscores it on this Index.
- Index
- 48
- Context
- 1M
- In / 1M
- —
- Out / 1M
- —
- License
- —
- Released
- 19 Feb 2026
AA lists cost per task, not a clean input/output pair we independently confirmed on Google's pricing page.
- Grok 4.7
xAI · Frontier closed
xAI's current coding and long-running agent flagship; same list price as Grok 4.6. AA Index 46 on v4.3.2 at xhigh.
- Index
- 46
- Context
- 500k
- In / 1M
- $2
- Out / 1M
- $6
- License
- —
- Released
- 21 Sep 2026
Scored: xhigh
Same standard band as Grok 4.6 under 200k prompt tokens ($2 / $0.50 cached / $6). Over 200k doubles. Fast variant is 2×. Score is Index v4.3.2 — not directly comparable to v4.1.1 rows on this board.
- Gemini 3.8 Flash
Google · Frontier closed
Google's Sep 2026 Flash flagship for long-horizon agents and coding; AA Index 41 on v4.3 at high thinking.
- Index
- 41
- Context
- 1.05M
- In / 1M
- $0.75
- Out / 1M
- $3.75
- License
- —
- Released
- 2 Sep 2026
Scored: high
Introductory Google list through 31 Dec 2026; then $1.50 / $7.50 from 1 Jan 2027. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board.
- DeepSeek V4.1 Flash
DeepSeek · Open weight
552B MoE (16B active), MIT weights; DeepSeek’s Sep 2026 Flash cut. Strong coding/agent claims; AA Index 40 on v4.3.
- Index
- 40
- Context
- 1M
- In / 1M
- $0.30
- Out / 1M
- $1.20
- License
- MIT
- Released
- 10 Sep 2026
Scored: max effort
First-party API as measured by Artificial Analysis. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board.
- Qwen3.5 397B A17B Int4
Alibaba / Qwen · Open weight
GPTQ Int4 open-weight cut of Qwen3.5 397B (397B MoE / 17B active) for cheaper local serving; AA Index is from the base reasoning model, not a separate int4 eval.
- Index
- 34
- Context
- 262k
- In / 1M
- —
- Out / 1M
- —
- License
- Apache 2.0
- Released
- February 2026
Scored: AA scored Qwen3.5 397B A17B (Reasoning); no separate int4 eval
Self-host GPTQ Int4 weights (no first-party int4 token price). Alibaba Cloud API for the base family is $0.60/$3.60 per 1M.
- Claude Haiku 4.5
Anthropic · Frontier closed
Fast, cheap Claude still on the current Anthropic price card (knowledge cutoff Feb 2025 on that page).
- Index
- 30
- Context
- 200k
- In / 1M
- $1
- Out / 1M
- $5
- License
- —
- Released
- 15 Oct 2025
- Gemma 4 31B
Google · Open weight
Google's current Gemma dense open-weight; Apache 2.0, 256k context.
- Index
- 30
- Context
- 256k
- In / 1M
- —
- Out / 1M
- —
- License
- Apache 2.0
- Released
- 2 Apr 2026
AA lists $0 / $0; treat as self-host, not a Google list price.
- Mistral Medium 3.5
Mistral · Open weight
Mistral's current mid-size open-weight; Large 3 scores well below this on the Index.
- Index
- 30
- Context
- 256k
- In / 1M
- $1.50
- Out / 1M
- $7.50
- License
- Modified MIT
- Released
- 29 Apr 2026
- Llama 4 Maverick
Meta · Open weight
Still the Llama-branded open-weight you can download; far behind Muse Spark 1.3 on current Meta API quality.
- Index
- 14
- Context
- 1M
- In / 1M
- $0.26
- Out / 1M
- $0.91
- License
- Llama 4 Community License
- Released
- 5 Apr 2025
Median provider price on AA, not a first-party Meta API.
| Lab | Kind | License | ||||||
|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 Anthropic's Sep 2026 coding and research flagship; tops the current AA index. Scored: adaptive reasoning, max effort, default fallback Cache reads $0.25 per 1M tokens (0.025× input). Input/output list price unchanged from Fable 5. | Anthropic | Frontier closed | 66 | 1M | $10 | $50 | — | 1 Sep 2026 |
| Claude Mythos 5.1 Invite-only (not publicly available); same weights as Fable 5.1; more permissive cyber/life-sciences safeguards via Project Glasswing / trusted access. Scored: AA scored Fable 5.1 (identical weights); adaptive reasoning, max effort, default fallback Not public API access — Project Glasswing / trusted access only. Cache reads $0.25 per 1M tokens. List price matches Fable 5.1 where invited. | Anthropic | Frontier closed | 66 | 1M | $10 | $50 | — | 1 Sep 2026 |
| Claude Opus 5 Anthropic's current flagship for long-horizon agents and deep reasoning. Scored: adaptive reasoning, max effort | Anthropic | Frontier closed | 63 | 1M | $5 | $25 | — | 24 Jul 2026 |
| Claude Fable 5 Highest-priced Claude in the current lineup; reserved for the heaviest agent jobs. Scored: adaptive reasoning, max effort, Opus 4.8 fallback | Anthropic | Frontier closed | 62 | 1M | $10 | $50 | — | 9 Jun 2026 |
| GPT-6 Astra OpenAI's GPT-6 flagship for hard end-to-end agent work; Critical-cyber tier with gated tools. Scored: max effort Short-context list. Long context bills $20 / $75 per 1M. Cache hit $1.00 / $2.00 short/long. Rolling out via Trusted Access first. | OpenAI | Frontier closed | 61 | 1.05M | $10 | $50 | — | 3 Sep 2026 |
| GPT-5.6 Sol OpenAI's GPT-5.6 flagship; the gpt-5.6 alias routes here. Scored: max effort Promotional list price through at least 21 Nov 2026. Prompts over 272k input tokens bill 2× input and 1.5× output. | OpenAI | Frontier closed | 61 | 1.05M | $4 | $20 | — | 9 Jul 2026 |
| Grok 4.6 Prior xAI coding/agent flagship on a 500k window; prefer Grok 4.7 for the current cut. Scored: high effort Standard band under 200k prompt tokens, per xAI docs. | xAI | Frontier closed | 61 | 500k | $2 | $6 | — | 12 Aug 2026 |
| Kimi K3 Highest-ranked open-weight model on the current Intelligence Index (2.8T MoE, 104B active). Scored: max effort | Moonshot / Kimi | Open weight | 60 | 1.05M | $3 | $15 | Kimi K3 License (commercial with restrictions) | 16 Jul 2026 |
| GLM-5.3 Zhipu's current open-weight frontier; ties Kimi K3 at 60 on the Index. Scored: max effort | Zhipu / Z.AI | Open weight | 60 | 1M | $1.40 | $4.40 | GLM-5 license | 14 Aug 2026 |
| Qwen3.8 2.4T A95B Alibaba's 2.4T MoE open-weight flagship (95B active, per the AA card). | Alibaba / Qwen | Open weight | 58 | 984k | $2 | $6 | Qwen3 license | 12 Aug 2026 |
| GLM-5.3-Flash Z.ai’s multimodal Flash cut of GLM-5.3 — 320B MoE / 18B active; ties GPT-5.6 Terra and Muse Spark at 57 on far cheaper tokens. Scored: max effort Native text+image input. ~10× cheaper list than GLM-5.3 ($1.40/$4.40). | Zhipu / Z.AI | Open weight | 57 | 1M | $0.15 | $0.50 | MIT | August 2026 |
| GPT-5.6 Terra Mid GPT-5.6 tier: most of Sol's window at half the input price. Scored: max effort | OpenAI | Frontier closed | 57 | 1.05M | $2 | $12 | — | 9 Jul 2026 |
| Muse Spark 1.2 Previous Meta API quality cut; kept for Index v4.1.1 comparison. Prefer Muse Spark 1.3 for the current Meta scorer. Scored: xhigh Prior Meta API scorer on Index v4.1.1; superseded by Muse Spark 1.3 for current Meta quality. | Meta | Frontier closed | 57 | 1.05M | $1.25 | $4.25 | — | 5 Aug 2026 |
| Gemini 3.7 Flash Google's current high-throughput Gemini; scores above 3.1 Pro Preview on this Index. Scored: high | Frontier closed | 56 | 1M | $0.75 | $3.75 | — | 13 Aug 2026 | |
| Claude Sonnet 5 Workhorse Claude: 1M context at Sonnet prices. Scored: adaptive reasoning, max effort | Anthropic | Frontier closed | 55 | 1M | $2 | $10 | — | 30 Jun 2026 |
| DeepSeek V4 Pro 0813 1.6T MoE, 49B active, MIT weights; the current DeepSeek quality cut. Scored: max effort First-party API as measured by Artificial Analysis (may differ from off-peak list). | DeepSeek | Open weight | 53 | 1M | $1.32 | $3.96 | MIT | 13 Aug 2026 |
| DeepSeek V4 Flash 0731 Cheaper DeepSeek V4 cut that still clears 50 on the Index. Scored: max effort First-party API as measured by Artificial Analysis. | DeepSeek | Open weight | 52 | 1M | $0.44 | $1.32 | MIT | 31 Jul 2026 |
| GPT-5.6 Luna Volume GPT-5.6 tier: same 1.05M window, priced for high-throughput work. Scored: max effort | OpenAI | Frontier closed | 52 | 1.05M | $0.20 | $1.20 | — | 9 Jul 2026 |
| Muse Spark 1.3 Meta's current API quality line (proprietary, not Llama weights); AA Index 48 on v4.3 at max reasoning. Scored: max Standard muse-spark-1.3 route (not used to improve products). Cached input $0.15 per 1M. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board. | Meta | Frontier closed | 48 | 1M | $1.25 | $4.25 | — | 2 Sep 2026 |
| Gemini 3.1 Pro Preview Google's Pro-class preview; Flash 3.7 now outscores it on this Index. AA lists cost per task, not a clean input/output pair we independently confirmed on Google's pricing page. | Frontier closed | 48 | 1M | — | — | — | 19 Feb 2026 | |
| Grok 4.7 xAI's current coding and long-running agent flagship; same list price as Grok 4.6. AA Index 46 on v4.3.2 at xhigh. Scored: xhigh Same standard band as Grok 4.6 under 200k prompt tokens ($2 / $0.50 cached / $6). Over 200k doubles. Fast variant is 2×. Score is Index v4.3.2 — not directly comparable to v4.1.1 rows on this board. | xAI | Frontier closed | 46 | 500k | $2 | $6 | — | 21 Sep 2026 |
| Gemini 3.8 Flash Google's Sep 2026 Flash flagship for long-horizon agents and coding; AA Index 41 on v4.3 at high thinking. Scored: high Introductory Google list through 31 Dec 2026; then $1.50 / $7.50 from 1 Jan 2027. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board. | Frontier closed | 41 | 1.05M | $0.75 | $3.75 | — | 2 Sep 2026 | |
| DeepSeek V4.1 Flash 552B MoE (16B active), MIT weights; DeepSeek’s Sep 2026 Flash cut. Strong coding/agent claims; AA Index 40 on v4.3. Scored: max effort First-party API as measured by Artificial Analysis. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board. | DeepSeek | Open weight | 40 | 1M | $0.30 | $1.20 | MIT | 10 Sep 2026 |
| Qwen3.5 397B A17B Int4 GPTQ Int4 open-weight cut of Qwen3.5 397B (397B MoE / 17B active) for cheaper local serving; AA Index is from the base reasoning model, not a separate int4 eval. Scored: AA scored Qwen3.5 397B A17B (Reasoning); no separate int4 eval Self-host GPTQ Int4 weights (no first-party int4 token price). Alibaba Cloud API for the base family is $0.60/$3.60 per 1M. | Alibaba / Qwen | Open weight | 34 | 262k | — | — | Apache 2.0 | February 2026 |
| Claude Haiku 4.5 Fast, cheap Claude still on the current Anthropic price card (knowledge cutoff Feb 2025 on that page). | Anthropic | Frontier closed | 30 | 200k | $1 | $5 | — | 15 Oct 2025 |
| Gemma 4 31B Google's current Gemma dense open-weight; Apache 2.0, 256k context. AA lists $0 / $0; treat as self-host, not a Google list price. | Open weight | 30 | 256k | — | — | Apache 2.0 | 2 Apr 2026 | |
| Mistral Medium 3.5 Mistral's current mid-size open-weight; Large 3 scores well below this on the Index. | Mistral | Open weight | 30 | 256k | $1.50 | $7.50 | Modified MIT | 29 Apr 2026 |
| Llama 4 Maverick Still the Llama-branded open-weight you can download; far behind Muse Spark 1.3 on current Meta API quality. Median provider price on AA, not a first-party Meta API. | Meta | Open weight | 14 | 1M | $0.26 | $0.91 | Llama 4 Community License | 5 Apr 2025 |
Sources
Figures as of 21 Sep 2026. They lag the labs. We fetched the pages below; we did not invent scores, windows, licenses, or list prices. Gemini 3.8 Flash list prices are from Google's Gemini API pricing page (introductory through 31 Dec 2026). Gemini 3.7 Flash dollars remain Artificial Analysis's Google API listing from an earlier pass. Gemini 3.1 Pro Preview has no token price on this table. Index v4.3 scores are not on the same scale as v4.1.1 rows.
- Artificial Analysis LLM Leaderboard — Intelligence Index scores, context windows, and first-party API input/output prices. Most rows still use Index v4.1.1 (as of early Sep 2026). Newly added/updated rows (Grok 4.7, Muse Spark 1.3, Gemini 3.8 Flash, DeepSeek V4.1 Flash) use Index v4.3 / v4.3.2 — do not treat those scores as on the same scale as v4.1.1 rows.
- Artificial Analysis Intelligence Index v4.1.1 — Composite of GDPval-AA v2, τ³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR. Index patched 6 Aug 2026.
- OpenAI — GPT-6 Astra — Flagship gpt-6-astra. 1,050,000 context. List $10 / $50 per 1M (short context); long context $20 / $75. Rolling out Sep 2026 via Trusted Access, then Plus/Pro/Business/Enterprise.
- OpenAI — GPT-5.6 Sol — Official context (1,050,000), list price $4 / $20 per 1M tokens, promotional through at least 21 Nov 2026.
- OpenAI — GPT-5.6 Terra — Official context and $2 / $12 per 1M tokens.
- OpenAI — GPT-5.6 Luna — Official context and $0.20 / $1.20 per 1M tokens.
- Anthropic — Claude Opus 5 overview — Released 24 Jul 2026. 1M context. Lineup prices: Fable $10/$50, Opus $5/$25, Sonnet $2/$10, Haiku 4.5 $1/$5.
- Anthropic — Claude Fable 5.1 and Mythos 5.1 — Announced 1 Sep 2026. Identical weights; Mythos is invite-only with a more permissive safeguard layer. API $10 / $50; cache reads $0.25 per 1M.
- Anthropic — Claude Mythos 5.1 overview — API id claude-mythos-5-1. Project Glasswing / trusted access. Same context and list price as Fable 5.1.
- Anthropic — API pricing — Claude Fable 5.1 and Mythos 5.1 list prices and cache multipliers.
- Artificial Analysis — Claude Fable 5.1 — Intelligence Index 66 (adaptive reasoning, max effort, default fallback). 1M context. $10 / $50. No separate Mythos 5.1 page; Mythos shares Fable 5.1 weights.
- Z.ai — GLM-5.3 — Announced 14 Aug 2026. Weights followed after a two-week safety pass.
- xAI — Introducing Grok 4.7 — Released 21 Sep 2026. Same list $2 / $6 as Grok 4.6 on the standard band; 500k context; coding/agent focus.
- Artificial Analysis — Grok 4.7 — Index 46 on AA Intelligence Index v4.3.2 (xhigh). In $2 / Out $6; 500k context; text+image in.
- xAI — Introducing Grok 4.6 — Released 12 Aug 2026.
- xAI — Grok 4.6 docs — 500,000 context. $2 / $6 per 1M tokens on the standard band.
- xAI API pricing — Confirms grok-4.6 flagship list price and 500k context.
- Artificial Analysis — GPT-6 Astra (max) — Index 61 at max effort; In $10 / Out $50; released Sep 2026. xhigh also 61; low 57.
- Artificial Analysis — GPT-5.6 Sol (max) — Index 61; In $4 / Out $20; released 9 Jul 2026.
- Artificial Analysis — Claude Opus 5 (max) — Index 63; In $5 / Out $25; released 24 Jul 2026.
- Artificial Analysis — Claude Fable 5 — Index 62; In $10 / Out $50; released 9 Jun 2026.
- Artificial Analysis — Claude Sonnet 5 (max) — Index 55; In $2 / Out $10; released 30 Jun 2026.
- Artificial Analysis — Kimi K3 (max) — Index 60; open weights; Kimi K3 License (commercial with restrictions); In $3 / Out $15; released 16 Jul 2026.
- Artificial Analysis — DeepSeek V4 Pro 0813 (max) — Index 53; MIT; In $1.32 / Out $3.96 on DeepSeek API as AA measured; released 13 Aug 2026.
- Artificial Analysis — GLM-5.3 (max) — Index 60 on the leaderboard; GLM-5 license; In $1.40 / Out $4.40.
- Z.ai — GLM-5.3-Flash — Multimodal Flash cut of GLM-5.3; 320B MoE / 18B active; MIT weights.
- Artificial Analysis — GLM-5.3-Flash — Index 57; In $0.15 / Out $0.50; 1M context; released August 2026; open weights MIT.
- Hugging Face — zai-org/GLM-5.3-Flash — Open weights download.
- Artificial Analysis — Qwen3.8 2.4T A95B — Index 58; Qwen3 license; In $2 / Out $6; released 12 Aug 2026.
- Artificial Analysis — Qwen3.5 397B A17B — Index 34 (reasoning). Alibaba API $0.60/$3.60; 262k context; February 2026. No separate int4 AA page.
- Hugging Face — Qwen3.5-397B-A17B-GPTQ-Int4 — GPTQ Int4 open weights; Apache 2.0.
- Artificial Analysis — Qwen3.5 397B A17B explainer — Background on the base 397B MoE / 17B active model.
- Artificial Analysis — Gemini 3.7 Flash (high) — Index 56; In $0.75 / Out $3.75 as AA lists Google API; August 2026.
- Artificial Analysis — Grok 4.6 — Index 61 at high effort; In $2 / Out $6.
- Artificial Analysis — Muse Spark 1.3 (max) — Index 48 on AA Intelligence Index v4.3.2 (max); In $1.25 / Out $4.25; 1M context; released 2 Sep 2026. Proprietary Meta API.
- Meta — Muse Spark 1.3 — Official Meta Model API: muse-spark-1.3 list $1.25 / $4.25; cached input $0.15; 1M context. Contributor route is cheaper and used to improve products.
- Artificial Analysis — Muse Spark 1.2 (xhigh) — Prior Meta API scorer; Index 57 on v4.1.1; In $1.25 / Out $4.25; August 2026. Kept as historical row.
- Artificial Analysis — Gemini 3.8 Flash (high) — Index 41 on AA Intelligence Index v4.3.2 (high); In $0.75 / Out $3.75; 1M context; released September 2026.
- Google — Gemini 3.8 Flash — Model id gemini-3.8-flash; 1,048,576 input context; thinking low/medium/high. Released September 2026.
- Google — Gemini API pricing (3.8 Flash) — Introductory $0.75 / $3.75 per 1M through 31 Dec 2026; then $1.50 / $7.50 from 1 Jan 2027.
- Artificial Analysis — Llama 4 Maverick — Open weights; Llama 4 Community License; Index 14; April 2025.
- Artificial Analysis — Gemma 4 31B — Apache 2; Index 30; April 2026. AA lists $0/$0 (self-host / no first-party list price).
- Artificial Analysis — Mistral Medium 3.5 — Modified MIT; Index 30; In $1.50 / Out $7.50; April 2026.
- Artificial Analysis — DeepSeek V4.1 Flash (max) — Index 40 on AA Intelligence Index v4.3 (max effort); MIT; In $0.30 / Out $1.20; 1M context; released 10 Sep 2026. 552B MoE / 16B active.
- Artificial Analysis — DeepSeek V4 Flash 0731 (max) — MIT; Index 52; In $0.44 / Out $1.32 as AA measured; July 2026.
