Artificial Analysis Intelligence Index
composite of several evaluations run by Artificial Analysis
Updated 11 min ago · first seen 11 Sept 2026
- Metric
- index ↑
- Current results
- 638
- Models
- 477
- Current leader
- Claude Fable 5.1 53.4
Score history · qwen3-4b-2507-instruct 1 row
Not enough history to chart — a single observation (6.74 on 11 Sept 2026). Rows under different configurations count separately; the list below shows each one.
- 6.74aa_slug=qwen3-4b-2507-instruct · version=4.3 · estimated=true11 Sept 2026
Frontier over time · index
8 leader changes recorded, all dated 11 Sept 2026 — the frontier line needs at least two distinct dates. The corpus is young: every result was first observed on the same day, so leader changes will separate in time as sources are re-crawled.
- 53.4Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 53.2Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 51.2Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 51.0gpt-6-astra OpenAI Independent11 Sept 2026
- 49.7gpt-6-astra OpenAI Independent11 Sept 2026
- 45.1Claude Opus 5 Anthropic Independent11 Sept 2026
- 14.6grok-3-mini-reasoning SpaceXAI Independent11 Sept 2026
- 5.71llama-2-chat-7b Meta AI Independent11 Sept 2026
Includes closed rows (history). A point is emitted whenever a result beats every earlier result of the same group, ordered by evaluated_at when the source publishes it, else observed_at.
Leaderboard 477 models
Select models with +, then Compare.
| # | Model | Score | Trust | Configuration | vs leader | Evaluated | Source | Actions |
|---|---|---|---|---|---|---|---|---|
| 201 | Qwen3 VL 32B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 11.9 | Independent | reasoningonversion4.3 | Partially comparable0.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 201 | ling-3-0-tinyOpen weightsinclusionAI | 11.9 | Independent | version4.3 | Comparable0.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 203 | Magistral Medium 1.2ClosedMistral AI · Magistral | 11.8 | Independent | version4.3 | Comparable-0.04 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 204 | Sonar Reasoning ProClosedPerplexity AI · Sonar | 11.8 | Independent | version4.3 | Comparable-0.05 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 205 | Granite 4.2 8BOpen weightsIBM · Granite 4.2 | 11.8 | Independent | version4.3 | Comparable-0.06 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 206 | hypernova-60bOpen weightsMultiverse Computing | 11.7 | Independent | version4.3 | Comparable-0.13 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 207 | MiniMax-M1-80kOpen weightsMiniMax · MiniMax | 11.7 | Independent | version4.3 | Comparable-0.14 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 208 | nemotron-cascade-2-30b-a3bOpen weightsNVIDIA · Nemotron | 11.7 | Independent | version4.3 | Comparable-0.20 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 209 | gemini-2-5-flash-reasoning-04-2025ClosedGoogle · Gemini 2.5 | 11.7 | Independent | version4.3 | Comparable-0.22 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 210 | Mercury 2ClosedInception | 11.5 | Independent | version4.3 | Comparable-0.36 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 211 | k2-think-v2Open weightsMBZUAI Institute of Foundation Models | 11.5 | Independent | version4.3 | Comparable-0.37 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 212 | longcat-flash-liteOpen weightsLongCat | 11.5 | Independent | version4.3 | Comparable-0.40 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 213 | Mistral Small 4Open weightsMistral AI · Mistral · best of 2 rows | 11.4 | Independent | version4.3 | Comparable-0.42 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 214 | deepseek-r1-0120Open weightsDeepSeek · DeepSeek | 11.4 | Independent | version4.3 | Comparable-0.46 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 215 | o1 PreviewClosedOpenAI · OpenAI o-series | 11.4 | Independent | version4.3 | Comparable-0.49 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 216 | hyperclova-x-seed-think-32bOpen weightsNaver · Seed | 11.4 | Independent | version4.3 | Comparable-0.51 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 217 | GLM 4.6VOpen weightsZ.ai (Zhipu AI) · GLM4.6 · best of 2 rows | 11.2 | Independent | version4.3 | Comparable-0.65 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 218 | Qwen3 Next 80B A3B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 11.2 | Independent | reasoningonversion4.3 | Partially comparable-0.67 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 219 | GLM 4.5 AirOpen weightsZ.ai (Zhipu AI) · GLM4.5 | 11.1 | Independent | version4.3 | Comparable-0.78 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 220 | mi-dm-k-2-5-pro-dec28ClosedKorea Telecom | 11.0 | Independent | version4.3 | Comparable-0.83 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 221 | ring-1tOpen weightsinclusionAI | 10.9 | Independent | version4.3 | Comparable-0.97 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 222 | Trinity Large ThinkingOpen weightsArcee AI | 10.9 | Independent | version4.3 | Comparable-0.99 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 223 | g9v3-3bOpen weightsAI9Stars | 10.8 | Independent | version4.3 | Comparable-1.03 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 224 | intellect-3Open weightsPrime Intellect | 10.6 | Independent | version4.3 | Comparable-1.26 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 225 | gpt-5-chatgptClosedOpenAI · GPT 5 | 10.4 | Independent | version4.3 | Comparable-1.43 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 226 | solar-open-100b-reasoningOpen weightsUpstage · Solar | 10.4 | Independent | version4.3 | Comparable-1.51 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 227 | gemini-2-5-flash-lite-preview-09-2025ClosedGoogle · Gemini 2.5 · best of 2 rows | 10.3 | Independent | reasoningonversion4.3 | Partially comparable-1.52 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 228 | nemotron-3-nano-omni-30b-a3bOpen weightsNVIDIA · Nemotron 3 | 10.3 | Independent | version4.3 | Comparable-1.62 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 229 | gpt-4.1-miniClosedOpenAI · GPT 4.1 | 10.2 | Independent | version4.3 | Comparable-1.71 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 230 | Qwen3 Coder NextOpen weightsQwen · Qwen3 | 10.1 | Independent | version4.3 | Comparable-1.82 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 231 | MiniMax-M1-40kOpen weightsMiniMax · MiniMax | 9.99 | Independent | version4.3 | Comparable-1.88 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 232 | gpt-oss-20bOpen weightsOpenAI · gpt-oss · best of 2 rows | 9.95 | Independent | reasoning_effortlowversion4.3 | Partially comparable-1.92 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 233 | Mistral Medium 3.1ClosedMistral AI · Mistral | 9.89 | Independent | version4.3 | Comparable-1.98 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 234 | k2-v2Open weightsMBZUAI Institute of Foundation Models · best of 3 rows | 9.87 | Independent | version4.3 | Comparable-2.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 235 | qwen3-30b-a3b-2507Open weightsAlibaba Group · Qwen3 · best of 2 rows | 9.81 | Independent | reasoningonversion4.3 | Partially comparable-2.06 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 236 | o1-miniClosedOpenAI · OpenAI o-series | 9.77 | Independent | version4.3 | Comparable-2.10 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 237 | DeepSeek V3 0324Open weightsDeepSeek · DeepSeek-V3 | 9.72 | Independent | version4.3 | Comparable-2.15 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 238 | Mistral Large 3Open weightsMistral AI · Mistral | 9.71 | Independent | version4.3 | Comparable-2.16 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 239 | ling-2-6-flashOpen weightsinclusionAI | 9.68 | Independent | version4.3 | Comparable-2.19 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 240 | Qwen3 Coder 30B A3B InstructOpen weightsQwen · Qwen3 | 9.59 | Independent | version4.3 | Comparable-2.28 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 240 | tri-21b-think-previewOpen weightsTrillion Labs | 9.59 | Independent | version4.3 | Comparable-2.28 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 242 | GPT-4.5 PreviewClosedOpenAI · GPT 4.5 | 9.58 | Independent | version4.3 | Comparable-2.29 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 243 | diffusiongemma-26b-a4bOpen weightsGoogle | 9.51 | Independent | version4.3 | Comparable-2.36 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 244 | qwen3-235b-a22b-instructOpen weightsAlibaba Group · Qwen3 · best of 2 rows | 9.50 | Independent | reasoningonversion4.3 | Partially comparable-2.37 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 245 | QwQ-32BOpen weightsAlibaba Group · Qwen | 9.47 | Independent | version4.3 | Comparable-2.40 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 246 | Qwen3 VL 30B A3B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 9.45 | Independent | reasoningonversion4.3 | Partially comparable-2.42 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 247 | gemini-2.0-flash-thinking-exp-01-21ClosedGoogle · Gemini 2.0 | 9.42 | Independent | version4.3 | Comparable-2.45 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 248 | Devstral 2Open weightsMistral AI · Devstral 2 | 9.41 | Independent | version4.3 | Comparable-2.46 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 249 | Llama 4 MaverickOpen weightsMeta AI · Llama 4 | 9.30 | Independent | version4.3 | Comparable-2.57 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 250 | motif-2-12-7bClosedMotif Technologies | 9.19 | Independent | version4.3 | Comparable-2.68 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 251 | ling-1tOpen weightsinclusionAI | 9.17 | Independent | version4.3 | Comparable-2.70 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 252 | nova-premierClosedAmazon Web Services · Nova | 9.16 | Independent | version4.3 | Comparable-2.71 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 253 | solar-pro-2-previewClosedUpstage · Solar · best of 2 rows | 9.07 | Independent | reasoningonversion4.3 | Partially comparable-2.80 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 254 | granite-4.2-3bOpen weightsIBM · Granite 4.2 | 9.06 | Independent | version4.3 | Comparable-2.81 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 255 | Mistral Medium 3ClosedMistral AI · Mistral | 9.05 | Independent | version4.3 | Comparable-2.82 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 255 | magistral-mediumClosedMistral AI · Magistral | 9.05 | Independent | version4.3 | Comparable-2.82 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 257 | devstral-mediumClosedMistral AI · Devstral | 9.01 | Independent | version4.3 | Comparable-2.86 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 257 | llama-nemotron-super-49b-v1-5Open weightsNVIDIA · Llama · best of 2 rows | 9.01 | Independent | reasoningonversion4.3 | Partially comparable-2.86 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 259 | tri-21b-think-v0-5Open weightsTrillion Labs | 8.99 | Independent | version4.3 | Comparable-2.88 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 260 | gpt-4o-chatgpt-03-25ClosedOpenAI · GPT 4 | 8.96 | Independent | version4.3 | Comparable-2.91 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 261 | Claude Haiku 3.5ClosedAnthropic · Claude | 8.94 | Independent | version4.3 | Comparable-2.93 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 261 | Gemini 2.0 FlashClosedGoogle · Gemini | 8.94 | Independent | version4.3 | Comparable-2.93 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 263 | llama-3-3-nemotron-super-49bOpen weightsNVIDIA · Llama 3.3 · best of 2 rows | 8.93 | Independent | reasoningonversion4.3 | Partially comparable-2.94 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 264 | gemma-4-E4BOpen weightsGoogle · Gemma 4 · best of 2 rows | 8.91 | Independent | version4.3 | Comparable-2.96 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 265 | Nemotron 3 Nano 30B A3BOpen weightsNVIDIA · Nemotron 3 · best of 2 rows | 8.90 | Independent | reasoningonversion4.3 | Partially comparable-2.97 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 266 | minicpm5-1bOpen weightsOpenBMB · best of 2 rows | 8.80 | Independent | version4.3 | Comparable-3.07 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 266 | qwen3-4b-2507-instructOpen weightsAlibaba Group · Qwen3 · best of 2 rows | 8.80 | Independent | reasoningonversion4.3 | Partially comparable-3.07 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 268 | sarvam-105bOpen weightsSarvam | 8.79 | Independent | version4.3 | Comparable-3.08 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 269 | gemini-2-0-pro-experimental-02-05ClosedGoogle · Gemini 2.0 | 8.75 | Independent | version4.3 | Comparable-3.12 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 270 | Devstral Small 1.0Open weightsMistral AI · Devstral | 8.74 | Independent | version4.3 | Comparable-3.13 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 271 | claude-3-opusClosedAnthropic · Claude 3 | 8.72 | Independent | version4.3 | Comparable-3.15 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 272 | SonarClosedPerplexity AI · Sonar · best of 2 rows | 8.67 | Independent | reasoningonversion4.3 | Partially comparable-3.20 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 273 | gemini-2-5-flash-04-2025ClosedGoogle · Gemini 2.5 | 8.66 | Independent | version4.3 | Comparable-3.21 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 274 | Magistral Small 1.2Open weightsMistral AI · Magistral | 8.59 | Independent | version4.3 | Comparable-3.28 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 275 | Gemini 2.5 Flash-LiteClosedGoogle · Gemini · best of 2 rows | 8.55 | Independent | version4.3 | Comparable-3.32 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 276 | gpt-4oClosedOpenAI · GPT 4 | 8.44 | Independent | version4.3 | Comparable-3.43 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 277 | nanbeige4-1-3bOpen weightsNanbeige | 8.40 | Independent | version4.3 | Comparable-3.47 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 278 | LFM2.5-2.6B (free)Open weightsLiquid AI · LFM2.5 | 8.39 | Independent | version4.3 | Comparable-3.48 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 279 | DeepSeek-R1-Distill-Qwen-32BOpen weightsDeepSeek · Qwen | 8.38 | Independent | version4.3 | Comparable-3.49 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 280 | Devstral Small 2Open weightsMistral AI · Devstral | 8.36 | Independent | version4.3 | Comparable-3.51 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 281 | gemini-2.0-flash-expClosedGoogle · Gemini 2.0 | 8.22 | Independent | version4.3 | Comparable-3.65 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 281 | magistral-smallOpen weightsMistral AI · Magistral | 8.22 | Independent | version4.3 | Comparable-3.65 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 283 | exaone-4-0-32bOpen weightsLG AI Research · EXAONE 4.0 · best of 2 rows | 8.18 | Independent | reasoningonversion4.3 | Partially comparable-3.69 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 284 | Qwen3 VL 8B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 8.17 | Independent | reasoningonversion4.3 | Partially comparable-3.70 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 285 | DeepSeek-R1-0528-Qwen3-8BOpen weightsDeepSeek · Qwen3 | 8.08 | Independent | version4.3 | Comparable-3.79 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 286 | qwen-2-5-maxClosedAlibaba Group · Qwen | 8.02 | Independent | version4.3 | Comparable-3.85 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 287 | Hermes-4-70BRestricted weightsNous Research · Hermes 4 | 7.91 | Independent | version4.3 | Comparable-3.96 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 287 | gemini-1-5-proClosedGoogle · Gemini 1.5 | 7.91 | Independent | version4.3 | Comparable-3.96 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 289 | R1 Distill Llama 70BOpen weightsDeepSeek · Llama | 7.89 | Independent | version4.3 | Comparable-3.98 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 290 | claude-35-sonnetClosedAnthropic · Claude 35 | 7.88 | Independent | version4.3 | Comparable-3.99 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 291 | deepseek-r1-distill-qwen-14bOpen weightsDeepSeek · Qwen | 7.85 | Independent | version4.3 | Comparable-4.02 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 292 | falcon-h1r-7bOpen weightsTII UAE · Falcon | 7.83 | Independent | version4.3 | Comparable-4.04 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 293 | Solar Pro 3ClosedUpstage · Solar | 7.82 | Independent | version4.3 | Comparable-4.05 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 293 | gpt-4.1-nanoClosedOpenAI · GPT 4.1 | 7.82 | Independent | version4.3 | Comparable-4.05 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 295 | ling-flash-2-0Open weightsinclusionAI | 7.81 | Independent | version4.3 | Comparable-4.06 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 296 | gemma-4-E2BOpen weightsGoogle · Gemma 4 · best of 2 rows | 7.77 | Independent | version4.3 | Comparable-4.10 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 297 | qwen3-omni-30b-a3b-instructOpen weightsAlibaba Group · Qwen3 · best of 2 rows | 7.76 | Independent | reasoningonversion4.3 | Partially comparable-4.11 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 298 | GPT-4o (2024-08-06)ClosedOpenAI · GPT 4 | 7.74 | Independent | version4.3 | Comparable-4.13 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 299 | Qwen2.5 72B InstructOpen weightsQwen · Qwen2.5 | 7.73 | Independent | version4.3 | Comparable-4.14 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 300 | step-3-vl-10bOpen weightsStepFun · Step3 | 7.69 | Independent | version4.3 | Comparable-4.18 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
One row per canonical model — its best current row inside this comparability group (effort variants are folded into the model). Bars are relative to the page's best score. “vs leader” reads comparability: partially comparable = same task, conditions differ (reasoning effort, temperature, judge). Rules →