SI Score · updated Oct 8, 2026
Superintelligence leaderboard.
Open evidence. Clear rankings.
The latest frontier models, ranked by the SI Score — a composite of open benchmarks, scaled per benchmark and weighted across reasoning, coding, math and human preference. Every number links to its source and date, and each score carries a confidence % that rises as sources report.
The leaderboard
Full catalog →| 1 | Claude Fable 5.1 Anthropic | 80.2 | 100% confidence 100 percent, Full | $10.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 8, 2026 · factual citation Open source ↗ | — |
| 2 | Claude Opus 5.5 Anthropic | 78.4 | 100% confidence 100 percent, Full | $4.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 8, 2026 · factual citation Open source ↗ | — |
| 3 | GPT-6 Astra OpenAI | 77.6 | 100% confidence 100 percent, Full | $10.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 4 | Claude Fable 5 Anthropic | 76.8 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 5 | Claude Opus 5 Anthropic | 74.9 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 6 | GPT-6.1 Sol OpenAI | 73.0 | 100% confidence 100 percent, Full | $2.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 7 | GPT-5.6 Sol OpenAI | 72.1 | 100% confidence 100 percent, Full | not yet reported | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 8 | Kimi K3 Moonshot AI | 71.0 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 9 | Claude Opus 4.6 Anthropic | 70.5 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 10 | Muse Spark 1.3 Meta | 70.1 | 93% confidence 93 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 11 | Claude Opus 4.7 Anthropic | 70.1 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 12 | Gemini 3.7 Flash Google | 70.0 | 100% confidence 100 percent, Full | $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 8, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 13 | GPT-5.5 OpenAI | 69.5 | 100% confidence 100 percent, Full | not yet reported | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 14 | Claude Sonnet 5.5 Anthropic | 69.4 | 93% confidence 93 percent, High | $2.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 8, 2026 · factual citation Open source ↗ | — |
| 15 | GPT-6 Sol OpenAI | 68.9 | 100% confidence 100 percent, Full | not yet reported | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 16 | Gemini 3.8 Flash Google | 68.8 | 100% confidence 100 percent, Full | $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 8, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 17 | Gemini 3.5 Flash Google | 68.3 | 100% confidence 100 percent, Full | $1.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 8, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 18 | Claude Sonnet 4.6 Anthropic | 68.2 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 19 | GPT-5.6 Terra OpenAI | 68.2 | 100% confidence 100 percent, Full | not yet reported | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 20 | Gemini 3.1 Pro Preview Google | 68.1 | 100% confidence 100 percent, Full | $2.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 8, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 21 | MiniMax-M3 minimax | 68.0 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 22 | GLM-5.3 Z.ai | 67.7 | 93% confidence 93 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 23 | Claude Sonnet 5 Anthropic | 67.7 | 93% confidence 93 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 24 | Claude Opus 4.8 Anthropic | 67.5 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 25 | Grok 4.6 xAI | 66.6 | 100% confidence 100 percent, Full | not yet reported | 500Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 26 | Qwen3.6 Max Preview Alibaba / Qwen | 66.4 | 64% confidence 64 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 27 | DeepSeek V4.1 Flash DeepSeek | 66.2 | 69% confidence 69 percent, Medium | $0.15DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 28 | GLM-5.3-Flash Z.ai | 66.2 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 29 | Kimi K2 Thinking Turbo Moonshot AI | 66.0 | 64% confidence 64 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 30 | GLM-5.2 Z.ai | 65.7 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 31 | Gemini 3.6 Flash Google | 65.6 | 100% confidence 100 percent, Full | $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 8, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 32 | Grok 4.7 xAI | 65.5 | 100% confidence 100 percent, Full | $2.00xAI models & pricingPublished source fact
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 500KxAI models & pricingPublished source fact
Retrieved Oct 8, 2026 · factual citation Open source ↗ | — |
| 33 | Grok 4.5 xAI | 64.8 | 100% confidence 100 percent, Full | not yet reported | 500Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 34 | Claude Haiku 5.5 Anthropic | 64.8 | 100% confidence 100 percent, Full | $0.10Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 8, 2026 · factual citation Open source ↗ | — |
| 35 | Claude Opus 4.5 Anthropic | 64.6 | 100% confidence 100 percent, Full | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 36 | Kimi K2.6 Moonshot AI | 64.0 | 85% confidence 85 percent, High | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 37 | Qwen3.5 397B-A17B Alibaba / Qwen | 64.0 | 69% confidence 69 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 38 | GPT-5.5 Pro OpenAI | 63.5 | 53% confidence 53 percent, Medium | not yet reported | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 39 | GPT-5.6 Luna OpenAI | 63.4 | 100% confidence 100 percent, Full | not yet reported | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 40 | Gemma 4 26B A4B IT Google | 63.4 | 64% confidence 64 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 41 | GLM-5.1 Z.ai | 63.3 | 64% confidence 64 percent, Medium | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 42 | DeepSeek V4 Pro DeepSeek | 63.1 | 85% confidence 85 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 43 | Gemma 4 31B IT Google | 63.0 | 64% confidence 64 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 44 | DeepSeek V4 Pro 0813 DeepSeek | 63.0 | 69% confidence 69 percent, Medium | $0.66DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 45 | Qwen3.7 Plus Alibaba / Qwen | 62.7 | 64% confidence 64 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 46 | Qwen3.5 35B-A3B Alibaba / Qwen | 62.7 | 64% confidence 64 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 47 | MiMo-V2.5-Pro xiaomi | 62.6 | 88% confidence 88 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 48 | Muse Spark 1.2 Meta | 62.3 | 53% confidence 53 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 49 | Qwen3.8 Max Alibaba / Qwen | 62.0 | 53% confidence 53 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 50 | GPT-5.4 OpenAI | 61.9 | 69% confidence 69 percent, Medium | not yet reported | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 51 | Qwen3.8 27B Alibaba / Qwen | 61.9 | 69% confidence 69 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 52 | GPT-6 Luna OpenAI | 61.6 | 100% confidence 100 percent, Full | $0.10OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts
Retrieved Oct 8, 2026 · factual citation Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 53 | Qwen3.6 Plus Alibaba / Qwen | 61.6 | 80% confidence 80 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 54 | GLM-4.7 Z.ai | 61.3 | 69% confidence 69 percent, Medium | not yet reported | 205Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 55 | Muse Spark 1.1 Meta | 61.2 | 53% confidence 53 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 56 | DeepSeek V4 Flash 0731 DeepSeek | 61.0 | 69% confidence 69 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 57 | Hy3 tencent | 60.7 | 88% confidence 88 percent, High | not yet reported | 256Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 58 | Qwen3.5 Flash Alibaba / Qwen | 59.5 | 64% confidence 64 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 59 | GLM-5 Z.ai | 58.9 | 85% confidence 85 percent, High | not yet reported | 205Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 60 | Kimi K2.5 Moonshot AI | 58.8 | 100% confidence 100 percent, Full | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 61 | Qwen3.7 Max Alibaba / Qwen | 58.6 | 53% confidence 53 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 62 | GPT-5.5 Instant OpenAI | 58.3 | 64% confidence 64 percent, Medium | not yet reported | 400Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 63 | Claude Haiku 4.5 Anthropic | 58.1 | 64% confidence 64 percent, Medium | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 64 | MiniMax-M2.7 minimax | 57.9 | 88% confidence 88 percent, High | not yet reported | 205Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 65 | Grok 4.20 (Reasoning) xAI | 57.7 | 80% confidence 80 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 66 | MiMo-V2.6-Pro xiaomi | 57.4 | 88% confidence 88 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 67 | Gemini 3 Pro Preview Google | 56.8 | 53% confidence 53 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 68 | MiniMax-M2.5 minimax | 56.7 | 88% confidence 88 percent, High | not yet reported | 205Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 69 | GPT-5.4 nano OpenAI | 56.6 | 69% confidence 69 percent, Medium | not yet reported | 400Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 70 | MiMo-V2.5 xiaomi | 56.4 | 88% confidence 88 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 71 | DeepSeek V4 Flash DeepSeek | 56.2 | 53% confidence 53 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 72 | GPT-5.4 mini OpenAI | 55.8 | 69% confidence 69 percent, Medium | not yet reported | 400Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 73 | MiMo-V2.6-Flash xiaomi | 55.8 | 88% confidence 88 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 74 | MiniMax-M2 minimax | 55.7 | 88% confidence 88 percent, High | not yet reported | 205Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 75 | Qwen3.6 27B Alibaba / Qwen | 55.6 | 53% confidence 53 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 76 | Gemini 3.5 Flash Lite Google | 55.2 | 100% confidence 100 percent, Full | $0.30Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 8, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 77 | DeepSeek V3 0324 DeepSeek | 55.0 | 64% confidence 64 percent, Medium | not yet reported | 164Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 78 | Claude Opus 4 Anthropic | 54.8 | 96% confidence 96 percent, High | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 79 | Grok 4.3 xAI | 54.6 | 80% confidence 80 percent, High | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 80 | Gemini 3 Flash Preview Google | 54.4 | 85% confidence 85 percent, High | $0.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 8, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 81 | DeepSeek-V3 DeepSeek | 54.1 | 75% confidence 75 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 82 | GPT OSS 120B OpenAI | 53.9 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 83 | Claude Opus 4.1 Anthropic | 53.5 | 80% confidence 80 percent, High | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 84 | GLM-4.7-Flash Z.ai | 53.3 | 69% confidence 69 percent, Medium | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 85 | Step 3.5 Flash stepfun | 53.0 | 88% confidence 88 percent, High | not yet reported | 256Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 86 | Claude Sonnet 4.5 Anthropic | 52.7 | 80% confidence 80 percent, High | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 87 | Qwen3 30B A3B Alibaba / Qwen | 52.6 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 88 | GPT OSS 20B OpenAI | 52.5 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 89 | GPT-5.2 OpenAI | 51.4 | 53% confidence 53 percent, Medium | not yet reported | 400Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 90 | Claude Sonnet 3.5 v2 Anthropic | 51.3 | 75% confidence 75 percent, Medium | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 91 | Llama 4 Scout 17B Instruct Meta | 51.0 | 64% confidence 64 percent, Medium | not yet reported | 10Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 92 | GPT-5 OpenAI | 50.9 | 53% confidence 53 percent, Medium | not yet reported | 400Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 93 | Qwen3 32B Alibaba / Qwen | 50.4 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 94 | Qwen3 235B-A22B Alibaba / Qwen | 50.3 | 69% confidence 69 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 95 | Claude Sonnet 3.7 Anthropic | 49.8 | 80% confidence 80 percent, High | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 96 | GPT-4o (2024-05-13) OpenAI | 49.4 | 75% confidence 75 percent, Medium | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 97 | QwQ 32B Alibaba / Qwen | 48.3 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 98 | DeepSeek-R1 DeepSeek | 48.0 | 85% confidence 85 percent, High | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 99 | Llama-3.1-70B-Instruct Meta | 45.8 | 75% confidence 75 percent, Medium | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 100 | Gemini 2.5 Pro Google | 45.5 | 85% confidence 85 percent, High | $1.25Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 8, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 101 | Claude Sonnet 4 Anthropic | 45.2 | 96% confidence 96 percent, High | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 102 | Gemma 3 12B IT Google | 45.1 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 103 | o3-mini OpenAI | 43.5 | 75% confidence 75 percent, Medium | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 104 | Gemma 3 27B IT Google | 43.1 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 105 | Claude Haiku 3.5 Anthropic | 41.7 | 75% confidence 75 percent, Medium | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 106 | Claude Haiku 3 Anthropic | 39.9 | 75% confidence 75 percent, Medium | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 107 | Qwen2.5-Coder-32B-Instruct Alibaba / Qwen | 39.6 | 75% confidence 75 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 108 | Gemma 3 4B IT Google | 39.2 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 109 | GPT-4o (2024-08-06) OpenAI | 38.8 | 88% confidence 88 percent, High | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 110 | Llama-3.1-8B-Instruct Meta | 37.8 | 75% confidence 75 percent, Medium | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 111 | Mistral Large 2.1 mistral | 37.4 | 88% confidence 88 percent, High | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 112 | Llama-3.3-70B-Instruct Meta | 35.1 | 88% confidence 88 percent, High | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 113 | Mistral Medium 3 mistral | 33.0 | 85% confidence 85 percent, High | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | — |
| 114 | Llama 4 Maverick 17B Instruct Meta | 32.2 | 69% confidence 69 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
| 115 | Llama-3.2-1B Meta | 32.2 | 75% confidence 75 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 8, 2026 · MIT Open source ↗ | Open |
Ranked models appear first. Provisional models need more evidence and are available through the toggle (or the release catalog). Prices are USD per 1M tokens (official first-party API). Dotted values carry their source — hover or tap to see it. “not yet reported” means the source has not published that figure for this model.
Score against price, and over time
Best value models →Latest releases
All releases →Oct 7, 2026
Claude Haiku 5.5
Anthropic
SI 64.8 · confidence 100%
Oct 6, 2026
Mistral Large 4
mistral · provisional
SI 56.1 · confidence 32%
Oct 1, 2026
Grok Imagine Video 1.5 Lite
xAI · provisional
score not yet reported
Sep 29, 2026
Ling 3.1 Flash
inclusionai · provisional
score not yet reported
Sep 29, 2026
GPT-6.1 Sol
OpenAI
SI 73.0 · confidence 100%
Understand the numbers
Methodology
How the SI Score is built, what the confidence % means, which sources we use and their licenses.
What is superintelligence?
The research term, the corporate lane-name, and the new government spelling — disentangled and dated.
“Super Intelligence” in U.S. policy
Executive Order 14434 made it executive-branch vocabulary on Sep 29, 2026. What it says, verbatim.
Statements tracker
Who has adopted the term, who rejects it, and what comes next — dated, sourced, neutral.