The Aigentic logo
Subscribe

Models

The labs keep shipping. This page is the current flagship cut: closed APIs from OpenAI, Anthropic, Google, xAI, and Meta, plus the open-weight models that actually sit next to them on a shared harness — Kimi, GLM, Qwen, DeepSeek, Gemma, Mistral, Llama. Quality is the Artificial Analysis Intelligence Index (mostly v4.1.1; Muse Spark 1.3, Gemini 3.8 Flash, and DeepSeek V4.1 Flash on v4.3), not a press-release table.

Prices are USD per million tokens on the first-party (or AA-measured) API. If a field is missing, it is a dash, not a guess. Llama 4 is still here because that is the downloadable Meta line; Muse Spark 1.3 is the Meta API that currently scores on AA. We do not treat every 7B fine-tune as a frontier model.

Price vs valueAPI tiers: list $/1M next to what you get.Plan tiersChat/product subscriptions: Free → Enterprise.ArenaMultiple Arena boards as charts — Text, Code, Vision, Search, Agent.LiveBenchCategory charts + full table — contamination-resistant.Pricing historyInput and output $/1M over time — step chart.Usage AdvisorReprice your API mix; CSV stays in-browser.What ifTools: past stock buys.

Artificial Analysis Intelligence Index v4.1.1

Higher is better. Navy is frontier / API-only. Accent is open weight. Effort settings are the ones AA scored, not always the lab default. Bars are scaled to this board's score range (not from zero).

Frontier closedOpen weight
  • Claude Fable 5.166
  • Claude Mythos 5.166
  • Claude Opus 563
  • Claude Fable 562
  • GPT-6 Astra61
  • GPT-5.6 Sol61
  • Grok 4.661
  • Kimi K360
  • GLM-5.360
  • Qwen3.8 2.4T A95B58
  • GLM-5.3-Flash57
  • GPT-5.6 Terra57
  • Muse Spark 1.257
  • Gemini 3.7 Flash56
  • Claude Sonnet 555
  • DeepSeek V4 Pro 081353
  • DeepSeek V4 Flash 073152
  • GPT-5.6 Luna52
  • Muse Spark 1.348
  • Gemini 3.1 Pro Preview48
  • Grok 4.746
  • Gemini 3.8 Flash41
  • DeepSeek V4.1 Flash40
  • Qwen3.5 397B A17B Int434
  • Claude Haiku 4.530
  • Gemma 4 31B30
  • Mistral Medium 3.530
  • Llama 4 Maverick14

Filter

  • Claude Fable 5.1

    Anthropic · Frontier closed

    Anthropic's Sep 2026 coding and research flagship; tops the current AA index.

    Index
    66
    Context
    1M
    In / 1M
    $10
    Out / 1M
    $50
    License
    Released
    1 Sep 2026

    Scored: adaptive reasoning, max effort, default fallback

    Cache reads $0.25 per 1M tokens (0.025× input). Input/output list price unchanged from Fable 5.

  • Claude Mythos 5.1

    Anthropic · Frontier closed

    Invite-only (not publicly available); same weights as Fable 5.1; more permissive cyber/life-sciences safeguards via Project Glasswing / trusted access.

    Index
    66
    Context
    1M
    In / 1M
    $10
    Out / 1M
    $50
    License
    Released
    1 Sep 2026

    Scored: AA scored Fable 5.1 (identical weights); adaptive reasoning, max effort, default fallback

    Not public API access — Project Glasswing / trusted access only. Cache reads $0.25 per 1M tokens. List price matches Fable 5.1 where invited.

  • Claude Opus 5

    Anthropic · Frontier closed

    Anthropic's current flagship for long-horizon agents and deep reasoning.

    Index
    63
    Context
    1M
    In / 1M
    $5
    Out / 1M
    $25
    License
    Released
    24 Jul 2026

    Scored: adaptive reasoning, max effort

  • Claude Fable 5

    Anthropic · Frontier closed

    Highest-priced Claude in the current lineup; reserved for the heaviest agent jobs.

    Index
    62
    Context
    1M
    In / 1M
    $10
    Out / 1M
    $50
    License
    Released
    9 Jun 2026

    Scored: adaptive reasoning, max effort, Opus 4.8 fallback

  • GPT-6 Astra

    OpenAI · Frontier closed

    OpenAI's GPT-6 flagship for hard end-to-end agent work; Critical-cyber tier with gated tools.

    Index
    61
    Context
    1.05M
    In / 1M
    $10
    Out / 1M
    $50
    License
    Released
    3 Sep 2026

    Scored: max effort

    Short-context list. Long context bills $20 / $75 per 1M. Cache hit $1.00 / $2.00 short/long. Rolling out via Trusted Access first.

  • GPT-5.6 Sol

    OpenAI · Frontier closed

    OpenAI's GPT-5.6 flagship; the gpt-5.6 alias routes here.

    Index
    61
    Context
    1.05M
    In / 1M
    $4
    Out / 1M
    $20
    License
    Released
    9 Jul 2026

    Scored: max effort

    Promotional list price through at least 21 Nov 2026. Prompts over 272k input tokens bill 2× input and 1.5× output.

  • Grok 4.6

    xAI · Frontier closed

    Prior xAI coding/agent flagship on a 500k window; prefer Grok 4.7 for the current cut.

    Index
    61
    Context
    500k
    In / 1M
    $2
    Out / 1M
    $6
    License
    Released
    12 Aug 2026

    Scored: high effort

    Standard band under 200k prompt tokens, per xAI docs.

  • Kimi K3

    Moonshot / Kimi · Open weight

    Highest-ranked open-weight model on the current Intelligence Index (2.8T MoE, 104B active).

    Index
    60
    Context
    1.05M
    In / 1M
    $3
    Out / 1M
    $15
    License
    Kimi K3 License (commercial with restrictions)
    Released
    16 Jul 2026

    Scored: max effort

  • GLM-5.3

    Zhipu / Z.AI · Open weight

    Zhipu's current open-weight frontier; ties Kimi K3 at 60 on the Index.

    Index
    60
    Context
    1M
    In / 1M
    $1.40
    Out / 1M
    $4.40
    License
    GLM-5 license
    Released
    14 Aug 2026

    Scored: max effort

  • Qwen3.8 2.4T A95B

    Alibaba / Qwen · Open weight

    Alibaba's 2.4T MoE open-weight flagship (95B active, per the AA card).

    Index
    58
    Context
    984k
    In / 1M
    $2
    Out / 1M
    $6
    License
    Qwen3 license
    Released
    12 Aug 2026
  • GLM-5.3-Flash

    Zhipu / Z.AI · Open weight

    Z.ai’s multimodal Flash cut of GLM-5.3 — 320B MoE / 18B active; ties GPT-5.6 Terra and Muse Spark at 57 on far cheaper tokens.

    Index
    57
    Context
    1M
    In / 1M
    $0.15
    Out / 1M
    $0.50
    License
    MIT
    Released
    August 2026

    Scored: max effort

    Native text+image input. ~10× cheaper list than GLM-5.3 ($1.40/$4.40).

  • GPT-5.6 Terra

    OpenAI · Frontier closed

    Mid GPT-5.6 tier: most of Sol's window at half the input price.

    Index
    57
    Context
    1.05M
    In / 1M
    $2
    Out / 1M
    $12
    License
    Released
    9 Jul 2026

    Scored: max effort

  • Muse Spark 1.2

    Meta · Frontier closed

    Previous Meta API quality cut; kept for Index v4.1.1 comparison. Prefer Muse Spark 1.3 for the current Meta scorer.

    Index
    57
    Context
    1.05M
    In / 1M
    $1.25
    Out / 1M
    $4.25
    License
    Released
    5 Aug 2026

    Scored: xhigh

    Prior Meta API scorer on Index v4.1.1; superseded by Muse Spark 1.3 for current Meta quality.

  • Gemini 3.7 Flash

    Google · Frontier closed

    Google's current high-throughput Gemini; scores above 3.1 Pro Preview on this Index.

    Index
    56
    Context
    1M
    In / 1M
    $0.75
    Out / 1M
    $3.75
    License
    Released
    13 Aug 2026

    Scored: high

  • Claude Sonnet 5

    Anthropic · Frontier closed

    Workhorse Claude: 1M context at Sonnet prices.

    Index
    55
    Context
    1M
    In / 1M
    $2
    Out / 1M
    $10
    License
    Released
    30 Jun 2026

    Scored: adaptive reasoning, max effort

  • DeepSeek V4 Pro 0813

    DeepSeek · Open weight

    1.6T MoE, 49B active, MIT weights; the current DeepSeek quality cut.

    Index
    53
    Context
    1M
    In / 1M
    $1.32
    Out / 1M
    $3.96
    License
    MIT
    Released
    13 Aug 2026

    Scored: max effort

    First-party API as measured by Artificial Analysis (may differ from off-peak list).

  • DeepSeek V4 Flash 0731

    DeepSeek · Open weight

    Cheaper DeepSeek V4 cut that still clears 50 on the Index.

    Index
    52
    Context
    1M
    In / 1M
    $0.44
    Out / 1M
    $1.32
    License
    MIT
    Released
    31 Jul 2026

    Scored: max effort

    First-party API as measured by Artificial Analysis.

  • GPT-5.6 Luna

    OpenAI · Frontier closed

    Volume GPT-5.6 tier: same 1.05M window, priced for high-throughput work.

    Index
    52
    Context
    1.05M
    In / 1M
    $0.20
    Out / 1M
    $1.20
    License
    Released
    9 Jul 2026

    Scored: max effort

  • Muse Spark 1.3

    Meta · Frontier closed

    Meta's current API quality line (proprietary, not Llama weights); AA Index 48 on v4.3 at max reasoning.

    Index
    48
    Context
    1M
    In / 1M
    $1.25
    Out / 1M
    $4.25
    License
    Released
    2 Sep 2026

    Scored: max

    Standard muse-spark-1.3 route (not used to improve products). Cached input $0.15 per 1M. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board.

  • Gemini 3.1 Pro Preview

    Google · Frontier closed

    Google's Pro-class preview; Flash 3.7 now outscores it on this Index.

    Index
    48
    Context
    1M
    In / 1M
    Out / 1M
    License
    Released
    19 Feb 2026

    AA lists cost per task, not a clean input/output pair we independently confirmed on Google's pricing page.

  • Grok 4.7

    xAI · Frontier closed

    xAI's current coding and long-running agent flagship; same list price as Grok 4.6. AA Index 46 on v4.3.2 at xhigh.

    Index
    46
    Context
    500k
    In / 1M
    $2
    Out / 1M
    $6
    License
    Released
    21 Sep 2026

    Scored: xhigh

    Same standard band as Grok 4.6 under 200k prompt tokens ($2 / $0.50 cached / $6). Over 200k doubles. Fast variant is 2×. Score is Index v4.3.2 — not directly comparable to v4.1.1 rows on this board.

  • Gemini 3.8 Flash

    Google · Frontier closed

    Google's Sep 2026 Flash flagship for long-horizon agents and coding; AA Index 41 on v4.3 at high thinking.

    Index
    41
    Context
    1.05M
    In / 1M
    $0.75
    Out / 1M
    $3.75
    License
    Released
    2 Sep 2026

    Scored: high

    Introductory Google list through 31 Dec 2026; then $1.50 / $7.50 from 1 Jan 2027. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board.

  • DeepSeek V4.1 Flash

    DeepSeek · Open weight

    552B MoE (16B active), MIT weights; DeepSeek’s Sep 2026 Flash cut. Strong coding/agent claims; AA Index 40 on v4.3.

    Index
    40
    Context
    1M
    In / 1M
    $0.30
    Out / 1M
    $1.20
    License
    MIT
    Released
    10 Sep 2026

    Scored: max effort

    First-party API as measured by Artificial Analysis. Score is Index v4.3 — not directly comparable to v4.1.1 rows on this board.

  • Qwen3.5 397B A17B Int4

    Alibaba / Qwen · Open weight

    GPTQ Int4 open-weight cut of Qwen3.5 397B (397B MoE / 17B active) for cheaper local serving; AA Index is from the base reasoning model, not a separate int4 eval.

    Index
    34
    Context
    262k
    In / 1M
    Out / 1M
    License
    Apache 2.0
    Released
    February 2026

    Scored: AA scored Qwen3.5 397B A17B (Reasoning); no separate int4 eval

    Self-host GPTQ Int4 weights (no first-party int4 token price). Alibaba Cloud API for the base family is $0.60/$3.60 per 1M.

  • Claude Haiku 4.5

    Anthropic · Frontier closed

    Fast, cheap Claude still on the current Anthropic price card (knowledge cutoff Feb 2025 on that page).

    Index
    30
    Context
    200k
    In / 1M
    $1
    Out / 1M
    $5
    License
    Released
    15 Oct 2025
  • Gemma 4 31B

    Google · Open weight

    Google's current Gemma dense open-weight; Apache 2.0, 256k context.

    Index
    30
    Context
    256k
    In / 1M
    Out / 1M
    License
    Apache 2.0
    Released
    2 Apr 2026

    AA lists $0 / $0; treat as self-host, not a Google list price.

  • Mistral Medium 3.5

    Mistral · Open weight

    Mistral's current mid-size open-weight; Large 3 scores well below this on the Index.

    Index
    30
    Context
    256k
    In / 1M
    $1.50
    Out / 1M
    $7.50
    License
    Modified MIT
    Released
    29 Apr 2026
  • Llama 4 Maverick

    Meta · Open weight

    Still the Llama-branded open-weight you can download; far behind Muse Spark 1.3 on current Meta API quality.

    Index
    14
    Context
    1M
    In / 1M
    $0.26
    Out / 1M
    $0.91
    License
    Llama 4 Community License
    Released
    5 Apr 2025

    Median provider price on AA, not a first-party Meta API.

Sources

Figures as of 21 Sep 2026. They lag the labs. We fetched the pages below; we did not invent scores, windows, licenses, or list prices. Gemini 3.8 Flash list prices are from Google's Gemini API pricing page (introductory through 31 Dec 2026). Gemini 3.7 Flash dollars remain Artificial Analysis's Google API listing from an earlier pass. Gemini 3.1 Pro Preview has no token price on this table. Index v4.3 scores are not on the same scale as v4.1.1 rows.

the aigentic

Know what matters in AI.

The stories shaping AI, the tools worth trying, and what they mean for your work.

Morning Brief + Closing Time. Two emails every weekday.

Free. Unsubscribe anytime.