Artificial Analysis Intelligence Index
composite of several evaluations run by Artificial Analysis
Updated 29 min ago · first seen 11 Sept 2026
- Metric
- index ↑
- Current results
- 638
- Models
- 477
- Current leader
- Claude Fable 5.1 53.4
Score history · gpt-5 4 rows
Not enough history to chart — 4 observations, all dated 11 Sept 2026. Rows under different configurations count separately; the list below shows each one.
- 22.98aa_slug=gpt-5 · version=4.3 · estimated=true11 Sept 2026
- 22.87aa_slug=gpt-5-medium · version=4.3 · estimated=true · aa_variant_slug=gpt-5-medium11 Sept 2026
- 20.79aa_slug=gpt-5-low · version=4.3 · estimated=true · aa_variant_slug=gpt-5-low11 Sept 2026
- 11.45aa_slug=gpt-5-minimal · version=4.3 · estimated=true · aa_variant_slug=gpt-5-minimal11 Sept 2026
Frontier over time · index
8 leader changes recorded, all dated 11 Sept 2026 — the frontier line needs at least two distinct dates. The corpus is young: every result was first observed on the same day, so leader changes will separate in time as sources are re-crawled.
- 53.4Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 53.2Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 51.2Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 51.0gpt-6-astra OpenAI Independent11 Sept 2026
- 49.7gpt-6-astra OpenAI Independent11 Sept 2026
- 45.1Claude Opus 5 Anthropic Independent11 Sept 2026
- 14.6grok-3-mini-reasoning SpaceXAI Independent11 Sept 2026
- 5.71llama-2-chat-7b Meta AI Independent11 Sept 2026
Includes closed rows (history). A point is emitted whenever a result beats every earlier result of the same group, ordered by evaluated_at when the source publishes it, else observed_at.
Leaderboard 477 models
Select models with +, then Compare.
| # | Model | Score | Trust | Configuration | vs leader | Evaluated | Source | Actions |
|---|---|---|---|---|---|---|---|---|
| 101 | mimo-v2-0206Open weightsXiaomi | 22.4 | Independent | version4.3 | Comparable0.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 102 | MiMo-V2.5Open weightsXiaomi | 22.3 | Independent | version4.3 | Comparable-0.14 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 103 | GLM 4.7Open weightsZ.ai (Zhipu AI) · GLM4.7 · best of 2 rows | 22.2 | Independent | version4.3 | Comparable-0.20 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 104 | Kimi K2 ThinkingOpen weightsMoonshot AI · Kimi | 22.0 | Independent | version4.3 | Comparable-0.42 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 105 | Qwen3.6 27BOpen weightsQwen · Qwen3.6 · best of 2 rows | 21.9 | Independent | version4.3 | Comparable-0.53 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 106 | o3-proClosedOpenAI · OpenAI o-series | 21.9 | Independent | version4.3 | Comparable-0.57 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 107 | g9v3-39a5bOpen weightsAI9Stars | 21.8 | Independent | version4.3 | Comparable-0.61 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 108 | KAT-Coder-Pro V2ClosedKwaipilot | 21.7 | Independent | version4.3 | Comparable-0.74 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 109 | DeepSeek V3.2Open weightsDeepSeek · DeepSeek-V3 | 21.5 | Independent | version4.3 | Comparable-0.95 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 110 | Qwen3.5 397B A17BOpen weightsQwen · Qwen3.5 · best of 2 rows | 21.4 | Independent | reasoningoffversion4.3 | Partially comparable-1.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 111 | Qwen3 MaxClosedQwen · Qwen3 · best of 2 rows | 21.3 | Independent | reasoningonversion4.3 | Partially comparable-1.18 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 112 | gpt-5.4-nanoClosedOpenAI · GPT 5.4 · best of 3 rows | 21.2 | Independent | version4.3 | Comparable-1.20 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 113 | Claude Sonnet 4.5ClosedAnthropic · Claude · best of 2 rows | 21.2 | Independent | reasoningonversion4.3 | Partially comparable-1.26 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 114 | MiniMax M2.1Open weightsMiniMax · MiniMax | 20.9 | Independent | version4.3 | Comparable-1.49 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 115 | deepseek-v4-pro-0424-non-reasoningOpen weightsDeepSeek · DeepSeek | 20.8 | Independent | version4.3 | Comparable-1.61 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 116 | mimo-v2-flashOpen weightsXiaomi · best of 2 rows | 20.8 | Independent | reasoningonversion4.3 | Partially comparable-1.62 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 117 | Claude Opus 4ClosedAnthropic · Claude | 20.6 | Independent | version4.3 | Comparable-1.80 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 118 | gpt-5-miniClosedOpenAI · GPT 5 · best of 3 rows | 20.6 | Independent | reasoning_effortmediumversion4.3 | Partially comparable-1.84 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 119 | GPT-5.1-Codex MiniClosedOpenAI · GPT 5.1 | 20.4 | Independent | version4.3 | Comparable-2.06 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 119 | qwen3-5-omni-plusClosedAlibaba Group · Qwen3.5 | 20.4 | Independent | version4.3 | Comparable-2.06 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 121 | grok-4-1-fastClosedSpaceXAI · Grok 4.1 · best of 2 rows | 20.4 | Independent | reasoningonversion4.3 | Partially comparable-2.07 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 122 | o3ClosedOpenAI · OpenAI o-series | 20.2 | Independent | version4.3 | Comparable-2.24 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 123 | LongCat 2.0Open weightsMeituan | 19.7 | Independent | version4.3 | Comparable-2.75 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 123 | k-exaone-2-0-0803Open weightsLG AI Research · EXAONE 2.0 | 19.7 | Independent | version4.3 | Comparable-2.75 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 125 | Step 3.7 FlashOpen weightsStepFun · Step3.7 | 19.5 | Independent | version4.3 | Comparable-2.96 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 126 | Qwen3.5-35B-A3BOpen weightsQwen · Qwen3.5 · best of 2 rows | 19.3 | Independent | version4.3 | Comparable-3.11 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 127 | Claude Sonnet 4ClosedAnthropic · Claude · best of 2 rows | 18.9 | Independent | reasoningonversion4.3 | Partially comparable-3.52 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 128 | deepseek-v4-flash-0420-non-reasoningOpen weightsDeepSeek · DeepSeek | 18.9 | Independent | version4.3 | Comparable-3.56 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 129 | Qwen3.6 35B A3BOpen weightsQwen · Qwen3.6 · best of 2 rows | 18.8 | Independent | version4.3 | Comparable-3.63 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 130 | jt-35b-flashClosedChina Mobile | 18.7 | Independent | version4.3 | Comparable-3.78 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 131 | MiniMax M2Open weightsMiniMax · MiniMax | 18.6 | Independent | version4.3 | Comparable-3.81 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 132 | kat-coder-pro-v1ClosedKwaiKAT | 18.6 | Independent | version4.3 | Comparable-3.85 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 133 | GLM 4.6Open weightsZ.ai (Zhipu AI) · GLM4.6 · best of 2 rows | 18.5 | Independent | reasoningonversion4.3 | Partially comparable-3.91 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 134 | Muse Glimmer 30BOpen weightsMeta AI | 18.1 | Independent | version4.3 | Comparable-4.37 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 135 | grok-4-fastClosedSpaceXAI · Grok 4 · best of 2 rows | 17.9 | Independent | reasoningonversion4.3 | Partially comparable-4.50 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 136 | gemini-3-flashClosedGoogle · Gemini 3 | 17.9 | Independent | version4.3 | Comparable-4.51 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 137 | Qwen3.5-122B-A10BOpen weightsQwen · Qwen3.5 · best of 2 rows | 17.8 | Independent | reasoningoffversion4.3 | Partially comparable-4.69 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 138 | Claude 3.7 SonnetClosedAnthropic · Claude · best of 2 rows | 17.7 | Independent | reasoningonversion4.3 | Partially comparable-4.73 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 139 | Claude Haiku 4.5ClosedAnthropic · Claude · best of 2 rows | 17.6 | Independent | reasoningonversion4.3 | Partially comparable-4.85 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 140 | ring-2-6-1tOpen weightsinclusionAI | 17.3 | Independent | version4.3 | Comparable-5.16 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 141 | ling-2-6-1tOpen weightsinclusionAI | 17 | Independent | version4.3 | Comparable-5.44 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 142 | Step 3.5 FlashOpen weightsStepFun · Step3.5 · best of 2 rows | 17.0 | Independent | version4.3 | Comparable-5.48 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 143 | doubao-seed-codeClosedByteDance · Seed | 16.9 | Independent | version4.3 | Comparable-5.49 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 144 | Gemma 4 26B A4BOpen weightsGoogle · Gemma 4 · best of 2 rows | 16.7 | Independent | version4.3 | Comparable-5.77 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 145 | Gemini 2.5 ProClosedGoogle · Gemini 2.5 | 16.7 | Independent | version4.3 | Comparable-5.78 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 146 | o4-miniClosedOpenAI · OpenAI o-series | 16.6 | Independent | version4.3 | Comparable-5.79 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 147 | claude-4-opusClosedAnthropic · Claude 4 | 16.6 | Independent | version4.3 | Comparable-5.83 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 148 | DeepSeek V3.2 ExpOpen weightsDeepSeek · DeepSeek-V3 | 16.6 | Independent | version4.3 | Comparable-5.86 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 149 | qwen3-max-thinking-previewClosedAlibaba Group · Qwen3 | 16.3 | Independent | version4.3 | Comparable-6.15 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 150 | DeepSeek V3Open weightsDeepSeek · DeepSeek · best of 2 rows | 16.0 | Independent | version4.3 | Comparable-6.40 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 151 | Gemini 3.1 Flash-Lite PreviewClosedGoogle · Gemini 3.1 | 16.0 | Independent | version4.3 | Comparable-6.41 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 152 | gemini-2-5-flash-preview-09-2025ClosedGoogle · Gemini 2.5 · best of 2 rows | 15.5 | Independent | reasoningonversion4.3 | Partially comparable-6.97 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 153 | Gemma 4 31BOpen weightsGoogle · Gemma 4 · best of 2 rows | 15.4 | Independent | version4.3 | Comparable-7.02 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 154 | DeepSeek V3.1 TerminusOpen weightsDeepSeek · DeepSeek · best of 2 rows | 15.4 | Independent | reasoningonversion4.3 | Partially comparable-7.04 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 155 | Kimi K2 0905Open weightsMoonshot AI · Kimi | 15.3 | Independent | version4.3 | Comparable-7.15 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 156 | o1ClosedOpenAI · OpenAI o-series | 15.2 | Independent | version4.3 | Comparable-7.21 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 157 | gemini-2-5-pro-03-25ClosedGoogle · Gemini 2.5 | 15.0 | Independent | version4.3 | Comparable-7.48 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 158 | Mistral Medium 3.5Open weightsMistral AI · Mistral | 14.9 | Independent | version4.3 | Comparable-7.55 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 159 | GLM 4.7 FlashOpen weightsZ.ai (Zhipu AI) · GLM4.7 · best of 2 rows | 14.9 | Independent | version4.3 | Comparable-7.57 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 160 | granite-4.2-30bOpen weightsIBM · Granite 4.2 | 14.8 | Independent | version4.3 | Comparable-7.61 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 161 | grok-3-mini-reasoningClosedSpaceXAI · Grok 3 | 14.6 | Independent | version4.3 | Comparable-7.82 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 162 | gemini-2-5-pro-05-06ClosedGoogle · Gemini 2.5 | 14.5 | Independent | version4.3 | Comparable-7.92 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 163 | deepseek-v3-2-specialeOpen weightsDeepSeek · DeepSeek | 14.5 | Independent | version4.3 | Comparable-7.98 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 164 | k-exaoneOpen weightsLG AI Research · EXAONE · best of 2 rows | 14.4 | Independent | version4.3 | Comparable-8.07 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 165 | ernie-5-0-thinking-previewClosedBaidu · ERNIE 5.0 | 14.3 | Independent | version4.3 | Comparable-8.18 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 166 | grok-4.20-0309-non-reasoningClosedxAI · Grok | 14.2 | Independent | version4.3 | Comparable-8.24 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 167 | gemma-4-12BOpen weightsGoogle · Gemma 4 · best of 2 rows | 14.2 | Independent | version4.3 | Comparable-8.26 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 168 | nova-2-0-proClosedAmazon Web Services · Nova 2.0 · best of 3 rows | 14.2 | Independent | reasoningonreasoning_effortmediumversion4.3 | Partially comparable-8.28 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 169 | Grok Build 0.1ClosedxAI · Grok | 14.1 | Independent | version4.3 | Comparable-8.38 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 170 | command-a-plusOpen weightsCohere · Command | 13.9 | Independent | version4.3 | Comparable-8.52 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 171 | deepseek-v3-2-0925Open weightsDeepSeek · DeepSeek | 13.9 | Independent | version4.3 | Comparable-8.56 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 172 | apriel-v1-5-15b-thinkerOpen weightsServiceNow | 13.8 | Independent | version4.3 | Comparable-8.62 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 173 | DeepSeek V3.1Open weightsDeepSeek · DeepSeek-V3 · best of 2 rows | 13.7 | Independent | version4.3 | Comparable-8.73 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 174 | Qwen3.5-9BOpen weightsQwen · Qwen3.5 · best of 2 rows | 13.7 | Independent | version4.3 | Comparable-8.79 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 174 | nova-2-0-omniClosedAmazon Web Services · Nova 2.0 · best of 3 rows | 13.7 | Independent | reasoningonreasoning_effortmediumversion4.3 | Partially comparable-8.79 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 176 | NVIDIA Nemotron 3.5 Lightning 30B A3BOpen weightsNVIDIA · Nemotron 3.5 | 13.6 | Independent | version4.3 | Comparable-8.80 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 177 | Nemotron 3 SuperOpen weightsNVIDIA · Nemotron 3 | 13.6 | Independent | version4.3 | Comparable-8.89 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 178 | Qwen3 VL 235B A22B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 13.4 | Independent | reasoningonversion4.3 | Partially comparable-9.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 179 | apriel-v1-6-15b-thinkerOpen weightsServiceNow | 13.4 | Independent | version4.3 | Comparable-9.04 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 180 | nova-2-0-liteClosedAmazon Web Services · Nova 2.0 · best of 4 rows | 13.4 | Independent | reasoningonversion4.3 | Partially comparable-9.07 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 181 | exaone-4-5-33bOpen weightsLG AI Research · EXAONE 4.5 | 13.2 | Independent | version4.3 | Comparable-9.23 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 182 | minicpm5-2bOpen weightsOpenBMB | 13.1 | Independent | version4.3 | Comparable-9.30 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 183 | Qwen3.5-4BOpen weightsQwen · Qwen3.5 · best of 2 rows | 13.1 | Independent | version4.3 | Comparable-9.32 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 183 | deepseek-r1Open weightsDeepSeek · DeepSeek-R1 | 13.1 | Independent | version4.3 | Comparable-9.32 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 185 | Gemini 2.5 FlashClosedGoogle · Gemini 2.5 · best of 2 rows | 13.1 | Independent | reasoningonversion4.3 | Partially comparable-9.33 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 186 | gpt-5-nanoClosedOpenAI · GPT 5 · best of 3 rows | 13.0 | Independent | version4.3 | Comparable-9.45 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 187 | North Mini Code (free)Open weightsCohere | 12.8 | Independent | version4.3 | Comparable-9.60 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 188 | GLM 4.5Open weightsZ.ai (Zhipu AI) · GLM4.5 | 12.8 | Independent | version4.3 | Comparable-9.67 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 189 | Kimi K2 0711Open weightsMoonshot AI · Kimi | 12.7 | Independent | version4.3 | Comparable-9.72 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 190 | Qwen3 235B A22B Instruct 2507Open weightsQwen · Qwen3 · best of 2 rows | 12.7 | Independent | version4.3 | Comparable-9.73 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 191 | gpt-4.1ClosedOpenAI · GPT 4.1 | 12.7 | Independent | version4.3 | Comparable-9.75 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 192 | qwen3-max-previewClosedAlibaba Group · Qwen3 | 12.6 | Independent | version4.3 | Comparable-9.85 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 193 | o3-miniClosedOpenAI · OpenAI o-series · best of 2 rows | 12.5 | Independent | version4.3 | Comparable-9.97 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 193 | qwen3-5-omni-flashClosedAlibaba Group · Qwen3.5 | 12.5 | Independent | version4.3 | Comparable-9.97 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 195 | o1-proClosedOpenAI · OpenAI o-series | 12.4 | Independent | version4.3 | Comparable-10.0 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 196 | gpt-oss-120bOpen weightsOpenAI · gpt-oss · best of 2 rows | 12.3 | Independent | version4.3 | Comparable-10.1 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 197 | jt-miniClosedChina Mobile | 12.2 | Independent | version4.3 | Comparable-10.2 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 198 | grok-3ClosedSpaceXAI · Grok 3 · best of 2 rows | 12.1 | Independent | version4.3 | Comparable-10.3 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 199 | Seed-OSS-36B-InstructOpen weightsByteDance · Seed | 12.1 | Independent | version4.3 | Comparable-10.3 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 200 | qwen3-coder-480b-a35b-instructOpen weightsAlibaba Group · Qwen3-Coder | 11.9 | Independent | version4.3 | Comparable-10.5 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
One row per canonical model — its best current row inside this comparability group (effort variants are folded into the model). Bars are relative to the page's best score. “vs leader” reads comparability: partially comparable = same task, conditions differ (reasoning effort, temperature, judge). Rules →