The Latent Score: 2026-09-04
Release tli-2026-v4.1-2026-09-04c. Methodology tli-2026-v4.1.
As of 2026-09-04, Fable 5.1 leads The Latent Score on the point estimate at 128.2 points among 19 ranked models. Rankings describe published capability evidence, not a statistically proven winner or the best model for every workload.
This page preserves a specific publication. Its data does not change when a new release is published.
Published rankings
| Rank | Model and evidence | Score (points) | Stability range (points) | Contributing results |
|---|---|---|---|---|
| 1 | Fable 5.1 | 128.2 | 122.7–134.2 | 51 |
| 2 | Opus 5 | 121.6 | 114.9–126.6 | 78 |
| 3 | GPT-6 Astra | 121.4 | 110.6–146.7 | 32 |
| 4 | Fable 5 | 118.2 | 112.9–122.9 | 91 |
| 5 | Muse Spark 1.3 | 112.4 | 105.6–119.6 | 18 |
| 6 | GPT-5.6 Sol | 107.8 | 103.4–114.2 | 94 |
| 7 | Kimi K3 | 105.2 | 98.8–109.6 | 81 |
| 8 | GLM-5.3 | 103.7 | 95.3–108.9 | 50 |
| 9 | Muse Spark 1.2 | 102.4 | 93.3–107.5 | 45 |
| 10 | Gemini 3.8 Flash | 102.3 | 93.7–110.5 | 42 |
| 11 | Grok 4.6 | 101.4 | 96.2–108.2 | 52 |
| 12 | Qwen3.8-Max | 97.3 | 89.7–101.5 | 67 |
| 13 | Gemini 3.7 Flash | 94.7 | 85.5–104.4 | 53 |
| 14 | GPT-5.6 Terra | 93.9 | 89.6–101.3 | 67 |
| 15 | Sonnet 5 | 88.6 | 81.9–94.1 | 62 |
| 16 | DeepSeek V4 Pro | 87.8 | 81.2–97.1 | 50 |
| 17 | GPT-5.6 Luna | 86.2 | 82–93.5 | 63 |
| 18 | DeepSeek V4 Flash | 72.6 | 67.1–82.1 | 50 |
| 19 | Qwen3.8-27B | 70.7 | 61.7–80.9 | 48 |
Methodology and limitations in this release
The scale is anchored to mean 100, SD 15 across 18 frozen reference models. It is not human IQ or a percentage. 95 benchmarks contribute across 7 domains.
This older release did not preserve a written methodology in its public snapshot. Its recorded scores and scale are archived here; current methodology documentation must not be assumed to describe this historical calculation.
Reading this release
Compare model scores alongside their stability ranges and domain coverage. A higher point estimate does not establish superiority for every task. The linked model records disclose contributing and excluded results and their published configurations.
This release uses methodology version tli-2026-v4.1. Comparisons with another release should check both its methodology version and evidence date; a new data release does not necessarily mean the scoring method changed.
Citation and downloads
The Latent. “The Latent Score.” Release tli-2026-v4.1-2026-09-04c, dataset dated 2026-09-04. Permanent release page.
Download public scores with definitions (JSON) · Data usage terms