Latent Value & Efficiency

Compare model capability and cost.

A shortlist based on current scores and documented costs. The Value Pick is the cheapest option on the point-score frontier above our declared capability floor. Observed cost comes from Artificial Analysis, including reasoning tokens. Small score gaps do not establish a reliable capability advantage.

Scores combine the best published reasoning settings across benchmarks; cost describes one measured configuration and workload. This does not promise the same capability at that cost. Test shortlisted models on your workload before choosing.

Value Pick · starting point

GPT-5.6 Luna

The cheapest frontier model above the declared score floor. New evidence or a different workload can change the comparison.

Latent Score
81.3
Observed cost / task
$0.047
vs Sonnet 5
98% cheaper

+0.0 Latent points

Capability × observed cost

The Latent Frontier

Up and to the left is better. Filled red points are the buying ladder: no cheaper option has a higher current point score.

Latent Frontier: score against observed cost per taskModels farther up and to the left provide more Latent Score for a lower observed cost per benchmark task, as measured by Artificial Analysis. Each mark is score against cost; buying-ladder models are highlighted; hollow points are unranked.708498112126140$0.05$0.10$0.15$0.25$0.40$0.60$1$1.5$2.5$4GPT-6 Astra: score 139.2, $1.67/task · Buying ladderFable 5.1: score 134.7, $3.69/task · Better buy: GPT-6 AstraOpus 5: score 129.2, $2.34/task · Better buy: GPT-6 AstraFable 5: score 120, $3.14/task · Better buy: GPT-6 AstraMuse Spark 1.3: score 114.2, $0.55/task · Buying ladderGPT-5.6 Sol: score 111.9, $1.23/task · Better buy: Muse Spark 1.3Grok 4.6: score 105.7, $0.84/task · Better buy: Muse Spark 1.3Kimi K3: score 104, $0.84/task · Better buy: Muse Spark 1.3Gemini 3.8 Flash: score 102.8, $0.58/task · Better buy: Muse Spark 1.3Gemini 3.7 Flash: score 101.4, $0.40/task · Buying ladderGPT-5.6 Terra: score 95.1, $0.51/task · Better buy: Gemini 3.7 FlashQwen3.8-Max: score 93.1, $1.13/task · Better buy: Gemini 3.7 FlashMuse Spark 1.2: score 92.2, $0.40/task · Better buy: Gemini 3.7 FlashDeepSeek V4 Pro: score 86.2, $0.25/task · Buying ladderGPT-5.6 Luna: score 81.3, $0.047/task · Value PickSonnet 5: score 81.3, $2.29/task · Better buy: GPT-5.6 LunaDeepSeek V4 Flash: score 67.1, $0.11/task · Better buy: GPT-5.6 LunaGPT-6 AstraFable 5.1Opus 5Fable 5Muse Spark 1.3GPT-5.6 SolGrok 4.6Kimi K3Gemini 3.8 FlashGemini 3.7 FlashGPT-5.6 TerraQwen3.8-MaxMuse Spark 1.2DeepSeek V4 ProGPT-5.6 LunaSonnet 5DeepSeek V4 FlashObserved cost per Intelligence Index task · log scaleThe Latent Score
  • Buying ladder
  • Dominated

Opus 5.5, Gemini 4 Argon, Sonnet 5.5, MiMo-V2.6-Pro, GPT-6.1 Sol, GPT-6 Sol, DeepSeek V4.1-Flash, Grok 4.7, GLM-5.3, Claude Haiku 5.5, GPT-6 Luna, GLM 5.3 Flash, Mistral Large 4, Qwen3.8-27B has no Artificial Analysis cost measurement and cannot be plotted. Cost data: Artificial Analysis Intelligence Index v4.1.1, retrieved 2026-08-17.

Observed cost · Artificial Analysis

Compare higher-scoring options

The ladder follows current point scores and documented costs, cheapest first. The Value Pick is the cheapest frontier model clearing the declared Latent ≥ 80 floor. Small score differences do not establish reliable superiority.

  1. GPT-5.6 Luna

    OpenAI

    Value Pick

    Latent Score
    81.3
    Cost
    $0.047/task
    Audit
  2. +$0.20/task · +4.9 Latent points · 5.3× the cost

    DeepSeek V4 Pro

    DeepSeek

    Latent Score
    86.2
    Cost
    $0.25/task
    Audit
  3. +$0.15/task · +15.2 Latent points · 1.6× the cost

    Gemini 3.7 Flash

    Google

    Latent Score
    101.4
    Cost
    $0.40/task
    Audit
  4. +$0.15/task · +12.8 Latent points · 1.4× the cost

    Muse Spark 1.3

    Meta

    Latent Score
    114.2
    Cost
    $0.55/task
    Audit
  5. +$1.12/task · +25.0 Latent points · 3× the cost

    GPT-6 Astra

    OpenAI

    Highest current score

    Latent Score
    139.2
    Cost
    $1.67/task
    Audit

Why not these models?Each has an alternative with no higher cost and no lower current score.

  • GPT-5.6 Sol

    OpenAI

    Compare Muse Spark 1.3. It costs less and leads by 2.3 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    111.9
    Cost
    $1.23/task
    Audit
  • Fable 5.1

    Anthropic

    Compare GPT-6 Astra. It costs less and leads by 4.5 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    134.7
    Cost
    $3.69/task
    Audit
  • GPT-5.6 Terra

    OpenAI

    Compare Gemini 3.7 Flash. It costs less and leads by 6.3 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    95.1
    Cost
    $0.51/task
    Audit
  • Grok 4.6

    xAI

    Compare Muse Spark 1.3. It costs less and leads by 8.5 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    105.7
    Cost
    $0.84/task
    Audit
  • Muse Spark 1.2

    Meta

    Compare Gemini 3.7 Flash. It leads by 9.2 Latent points at the same cost. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    92.2
    Cost
    $0.40/task
    Audit
  • Opus 5

    Anthropic

    Compare GPT-6 Astra. It costs less and leads by 10.0 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    129.2
    Cost
    $2.34/task
    Audit
  • Kimi K3

    Moonshot AI

    Compare Muse Spark 1.3. It costs less and leads by 10.2 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    104.0
    Cost
    $0.84/task
    Audit
  • Gemini 3.8 Flash

    Google

    Compare Muse Spark 1.3. It costs less and leads by 11.4 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    102.8
    Cost
    $0.58/task
    Audit
  • DeepSeek V4 Flash

    DeepSeek

    Compare GPT-5.6 Luna. It costs less and leads by 14.2 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    67.1
    Cost
    $0.11/task
    Audit
  • Fable 5

    Anthropic

    Compare GPT-6 Astra. It costs less and leads by 19.2 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    120.0
    Cost
    $3.14/task
    Audit
  • Qwen3.8-Max

    Alibaba

    Compare Gemini 3.7 Flash. It costs less and leads by 8.3 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    93.1
    Cost
    $1.13/task
    Audit
  • Sonnet 5

    Anthropic

    Compare GPT-5.6 Luna. It costs less and leads by 0.0 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    81.3
    Cost
    $2.29/task
    Audit

Not ranked14 of 31 models

  • Opus 5.5Insufficient comparable evidence or no configuration-matched cost.144.2score-cost/task
  • Gemini 4 ArgonInsufficient comparable evidence or no configuration-matched cost.134.8score-cost/task
  • Sonnet 5.5Insufficient comparable evidence or no configuration-matched cost.132.5score-cost/task
  • MiMo-V2.6-ProInsufficient comparable evidence or no configuration-matched cost.125.8score-cost/task
  • GPT-6.1 SolInsufficient comparable evidence or no configuration-matched cost.118.2score-cost/task
  • GPT-6 SolInsufficient comparable evidence or no configuration-matched cost.117.5score-cost/task
  • DeepSeek V4.1-FlashInsufficient comparable evidence or no configuration-matched cost.113score-cost/task
  • Grok 4.7Insufficient comparable evidence or no configuration-matched cost.111.3score-cost/task
  • GLM-5.3Insufficient comparable evidence or no configuration-matched cost.103.8score-cost/task
  • Claude Haiku 5.5Insufficient comparable evidence or no configuration-matched cost.100.5score-cost/task
  • GPT-6 LunaInsufficient comparable evidence or no configuration-matched cost.97.3score-cost/task
  • GLM 5.3 FlashInsufficient comparable evidence or no configuration-matched cost.94.4score-cost/task
  • Mistral Large 4Insufficient comparable evidence or no configuration-matched cost.83score-cost/task
  • Qwen3.8-27BInsufficient comparable evidence or no configuration-matched cost.70.5score-cost/task

Secondary lens · listed price, fixed basket

Compare by listed token prices

Same purchasing rules, different definition of cost: 1,000,000 input + 250,000 output tokens at listed first-party rates. Reasoning-token consumption is invisible here. That is what the observed-cost view above adds.

  1. DeepSeek V4 Flash

    DeepSeek

    Latent Score
    67.1
    Cost
    $0.21 basket
    Audit
  2. +$0.01 basket · +33.4 Latent points · 1.1× the cost

    Claude Haiku 5.5

    Anthropic

    Value Pick

    Latent Score
    100.5
    Cost
    $0.23 basket
    Audit
  3. +$0.43 basket · +25.3 Latent points · 2.9× the cost

    MiMo-V2.6-Pro

    Xiaomi

    Latent Score
    125.8
    Cost
    $0.65 basket
    Audit
  4. +$3.85 basket · +9.0 Latent points · 6.9× the cost

    Gemini 4 Argon

    Google

    Latent Score
    134.8
    Cost
    $4.50 basket
    Audit
  5. +$4.50 basket · +9.4 Latent points · 2× the cost

    Opus 5.5

    Anthropic

    Highest current score

    Latent Score
    144.2
    Cost
    $9.00 basket
    Audit

Why not these models?Each has an alternative with no higher cost and no lower current score.

  • Sonnet 5.5

    Anthropic

    Compare Gemini 4 Argon. It leads by 2.3 Latent points at the same cost. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    132.5
    Cost
    $4.50 basket
    Audit
  • GPT-6 Luna

    OpenAI

    Compare Claude Haiku 5.5. It leads by 3.2 Latent points at the same cost. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    97.3
    Cost
    $0.23 basket
    Audit
  • GPT-6 Astra

    OpenAI

    Compare Opus 5.5. It costs less and leads by 5.0 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    139.2
    Cost
    $22.50 basket
    Audit
  • GLM 5.3 Flash

    Z.ai

    Compare Claude Haiku 5.5. It costs less and leads by 6.1 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    94.4
    Cost
    $0.28 basket
    Audit
  • Fable 5.1

    Anthropic

    Compare Gemini 4 Argon. It costs less and leads by 0.1 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    134.7
    Cost
    $22.50 basket
    Audit
  • Muse Spark 1.3

    Meta

    Compare MiMo-V2.6-Pro. It costs less and leads by 11.6 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    114.2
    Cost
    $2.31 basket
    Audit
  • Grok 4.7

    xAI

    Compare MiMo-V2.6-Pro. It costs less and leads by 14.5 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    111.3
    Cost
    $3.50 basket
    Audit
  • Opus 5

    Anthropic

    Compare Gemini 4 Argon. It costs less and leads by 5.6 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    129.2
    Cost
    $11.25 basket
    Audit
  • GPT-6.1 Sol

    OpenAI

    Compare MiMo-V2.6-Pro. It costs less and leads by 7.6 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    118.2
    Cost
    $4.50 basket
    Audit
  • GPT-6 Sol

    OpenAI

    Compare MiMo-V2.6-Pro. It costs less and leads by 8.3 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    117.5
    Cost
    $4.50 basket
    Audit
  • GPT-5.6 Luna

    OpenAI

    Compare Claude Haiku 5.5. It costs less and leads by 19.2 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    81.3
    Cost
    $0.50 basket
    Audit
  • Grok 4.6

    xAI

    Compare MiMo-V2.6-Pro. It costs less and leads by 20.1 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    105.7
    Cost
    $3.50 basket
    Audit
  • Gemini 3.8 Flash

    Google

    Compare MiMo-V2.6-Pro. It costs less and leads by 23.0 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    102.8
    Cost
    $1.69 basket
    Audit
  • Fable 5

    Anthropic

    Compare MiMo-V2.6-Pro. It costs less and leads by 5.8 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    120.0
    Cost
    $22.50 basket
    Audit
  • Gemini 3.7 Flash

    Google

    Compare MiMo-V2.6-Pro. It costs less and leads by 24.4 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    101.4
    Cost
    $1.69 basket
    Audit
  • Kimi K3

    Moonshot AI

    Compare MiMo-V2.6-Pro. It costs less and leads by 21.8 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    104.0
    Cost
    $6.75 basket
    Audit
  • GPT-5.6 Sol

    OpenAI

    Compare MiMo-V2.6-Pro. It costs less and leads by 13.9 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    111.9
    Cost
    $12.50 basket
    Audit
  • Qwen3.8-Max

    Alibaba

    Compare Claude Haiku 5.5. It costs less and leads by 7.4 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    93.1
    Cost
    $3.50 basket
    Audit
  • Muse Spark 1.2

    Meta

    Compare Claude Haiku 5.5. It costs less and leads by 8.3 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    92.2
    Cost
    $2.31 basket
    Audit
  • DeepSeek V4 Pro

    DeepSeek

    Compare Claude Haiku 5.5. It costs less and leads by 14.3 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    86.2
    Cost
    $2.31 basket
    Audit
  • GPT-5.6 Terra

    OpenAI

    Compare Claude Haiku 5.5. It costs less and leads by 5.4 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    95.1
    Cost
    $5.00 basket
    Audit
  • Mistral Large 4

    Mistral AI

    Compare Claude Haiku 5.5. It costs less and leads by 17.5 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    83.0
    Cost
    $2.40 basket
    Audit
  • Sonnet 5

    Anthropic

    Compare Claude Haiku 5.5. It costs less and leads by 19.2 Latent points. This point-score comparison does not establish a reliable capability advantage; evaluate both on the tasks and settings you plan to use.

    Latent Score
    81.3
    Cost
    $4.50 basket
    Audit

Not ranked3 of 31 models

  • DeepSeek V4.1-FlashInsufficient comparable evidence or no configuration-matched cost.113score-basket
  • GLM-5.3Insufficient comparable evidence or no configuration-matched cost.103.8score-basket
  • Qwen3.8-27BInsufficient comparable evidence or no configuration-matched cost.70.5score-basket

le-2026-v3.0 · lv-2026-v4.0

Rule-based purchasing criteria, two cost lenses

Full methodology →

What the cost measures

The headline view uses Artificial Analysis's measured cost to complete one Intelligence Index v4.1.1 task (input, cache, reasoning, and answer tokens included). Cost uses the recorded configuration; a missing or mismatched measurement means no cost ranking. Capability combines results across reasoning settings, so the score and measured cost do not describe one end-to-end operating point. Lower-effort score estimates from the earlier methodology are not included.

How the shortlist is ordered

The frontier contains models with no cheaper-or-equal alternative that has an equal-or-higher current score, with at least one strict advantage. This is a numerical comparison, not a finding that one model will perform better on your tasks. Placement follows point scores. Probability-of-superiority claims are not assessed in this methodology; consult each model’s interval and reporting sensitivity on the score page.

A declared capability floor

The Value Pick is the cheapest non-dominated model clearing Latent ≥ 80, so an extremely cheap but weak model can never headline on price alone. Declared capability floor of 80 on this version’s frozen 100±15 reference scale; a purchasing policy, not an empirical optimum.

A ladder, not a value race

Models above the pick are not "worse value" in some scalar sense. They are step-up options, ordered by cost, each stating the difference in current score alongside the extra cost. These differences are estimates, not guaranteed gains. The scalar Efficiency and Value scores published under the previous board versions remain in the released snapshots for reproducibility, but no displayed ordering uses them.

Observed cost per task measurements by Artificial Analysis (Intelligence Index v4.1.1 workload, retrieved 2026-08-17), used as cost evidence only. Latency, throughput, reliability, and human preference are not silently folded into any headline.