Artificial Analysis Intelligence Index
composite of several evaluations run by Artificial Analysis
Updated 28 min ago · first seen 11 Sept 2026
- Metric
- index ↑
- Current results
- 638
- Models
- 477
- Current leader
- Claude Fable 5.1 53.4
Score history · mistral-large-2 1 row
Not enough history to chart — a single observation (7.56 on 11 Sept 2026). Rows under different configurations count separately; the list below shows each one.
- 7.56aa_slug=mistral-large-2 · version=4.3 · estimated=true11 Sept 2026
Frontier over time · index
8 leader changes recorded, all dated 11 Sept 2026 — the frontier line needs at least two distinct dates. The corpus is young: every result was first observed on the same day, so leader changes will separate in time as sources are re-crawled.
- 53.4Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 53.2Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 51.2Claude Fable 5.1 Anthropic Independent11 Sept 2026
- 51.0gpt-6-astra OpenAI Independent11 Sept 2026
- 49.7gpt-6-astra OpenAI Independent11 Sept 2026
- 45.1Claude Opus 5 Anthropic Independent11 Sept 2026
- 14.6grok-3-mini-reasoning SpaceXAI Independent11 Sept 2026
- 5.71llama-2-chat-7b Meta AI Independent11 Sept 2026
Includes closed rows (history). A point is emitted whenever a result beats every earlier result of the same group, ordered by evaluated_at when the source publishes it, else observed_at.
Leaderboard 477 models
Select models with +, then Compare.
| # | Model | Score | Trust | Configuration | vs leader | Evaluated | Source | Actions |
|---|---|---|---|---|---|---|---|---|
| 301 | Llama 3.3 70BOpen weightsMeta AI · Llama 3.3 | 7.66 | Independent | version4.3 | Comparable0.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 302 | qwen3-30b-a3b-instructOpen weightsAlibaba Group · Qwen3 · best of 2 rows | 7.63 | Independent | reasoningonversion4.3 | Partially comparable-0.03 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 303 | Sonar ProClosedPerplexity AI · Sonar | 7.61 | Independent | version4.3 | Comparable-0.05 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 304 | devstral-smallOpen weightsMistral AI · Devstral | 7.60 | Independent | version4.3 | Comparable-0.06 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 305 | QwQ-32B-PreviewOpen weightsAlibaba Group · Qwen | 7.59 | Independent | version4.3 | Comparable-0.07 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 306 | GLM 4.5VOpen weightsZ.ai (Zhipu AI) · GLM4.5 · best of 2 rows | 7.56 | Independent | version4.3 | Comparable-0.10 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 306 | mistral-large-2Open weightsMistral AI · Mistral | 7.56 | Independent | version4.3 | Comparable-0.10 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 308 | llama-3-1-nemotron-ultra-253b-v1-reasoningOpen weightsNVIDIA · Llama 3.1 | 7.53 | Independent | version4.3 | Comparable-0.13 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 309 | ernie-4-5-300b-a47bOpen weightsBaidu · ERNIE 4.5 | 7.50 | Independent | version4.3 | Comparable-0.16 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 310 | Hermes 4 405BOpen weightsNous Research · Hermes 4 | 7.49 | Independent | version4.3 | Comparable-0.17 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 310 | solar-pro-2ClosedUpstage · Solar · best of 2 rows | 7.49 | Independent | reasoningonversion4.3 | Partially comparable-0.17 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 312 | nvidia-nemotron-nano-12b-v2-vlOpen weightsNVIDIA · Nemotron · best of 2 rows | 7.48 | Independent | reasoningonversion4.3 | Partially comparable-0.18 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 313 | granite-4.1-30bOpen weightsIBM · Granite 4.1 | 7.44 | Independent | version4.3 | Comparable-0.22 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 314 | NVIDIA-Nemotron-Nano-9B-v2Open weightsNVIDIA · Nemotron · best of 2 rows | 7.43 | Independent | version4.3 | Comparable-0.23 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 315 | gemini-2-0-flash-lite-001ClosedGoogle · Gemini 2.0 | 7.41 | Independent | version4.3 | Comparable-0.25 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 315 | hermes-4-llama-3-1-405bOpen weightsNous Research · Llama 3.1 | 7.41 | Independent | version4.3 | Comparable-0.25 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 317 | nvidia-nemotron-3-nano-4bOpen weightsNVIDIA · Nemotron 3 | 7.38 | Independent | version4.3 | Comparable-0.28 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 318 | Mistral Small 3.1Open weightsMistral AI · Mistral | 7.37 | Independent | version4.3 | Comparable-0.29 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 319 | qwen3-32b-instructOpen weightsAlibaba Group · Qwen3 | 7.34 | Independent | version4.3 | Comparable-0.32 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 320 | GPT-4o (2024-05-13)ClosedOpenAI · GPT 4 | 7.33 | Independent | version4.3 | Comparable-0.33 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 320 | gemini-2-0-flash-lite-previewClosedGoogle · Gemini 2.0 | 7.33 | Independent | version4.3 | Comparable-0.33 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 322 | llama-3-1-nemotron-nano-4b-reasoningOpen weightsNVIDIA · Llama 3.1 | 7.31 | Independent | version4.3 | Comparable-0.35 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 323 | Kimi-Linear-48B-A3B-InstructOpen weightsMoonshot AI · Kimi | 7.30 | Independent | version4.3 | Comparable-0.36 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 324 | Llama-3.1-405BRestricted weightsMeta AI · Llama 3.1 | 7.29 | Independent | version4.3 | Comparable-0.37 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 325 | Qwen3-4BOpen weightsQwen · Qwen3 | 7.23 | Independent | version4.3 | Comparable-0.43 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 326 | LFM2.5-8B-A1BOpen weightsLiquid AI · LFM2.5 | 7.22 | Independent | version4.3 | Comparable-0.44 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 326 | Qwen3 32BOpen weightsQwen · Qwen3 | 7.22 | Independent | version4.3 | Comparable-0.44 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 328 | claude-35-sonnet-june-24ClosedAnthropic · Claude 35 | 7.21 | Independent | version4.3 | Comparable-0.45 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 329 | tulu3-405bOpen weightsAllen Institute for AI | 7.20 | Independent | version4.3 | Comparable-0.46 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 330 | gpt-4o-chatgptClosedOpenAI · GPT 4 | 7.19 | Independent | version4.3 | Comparable-0.47 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 331 | Pixtral LargeOpen weightsMistral AI · Pixtral | 7.15 | Independent | version4.3 | Comparable-0.51 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 331 | ring-flash-2-0Open weightsinclusionAI | 7.15 | Independent | version4.3 | Comparable-0.51 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 333 | olmo-3-1-32b-instructOpen weightsAllen Institute for AI · OLMo 3.1 · best of 2 rows | 7.12 | Independent | reasoningonversion4.3 | Partially comparable-0.54 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 334 | grok-2Open weightsxAI · Grok 2 | 7.11 | Independent | version4.3 | Comparable-0.55 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 335 | gemini-1-5-flashClosedGoogle · Gemini 1.5 | 7.07 | Independent | version4.3 | Comparable-0.59 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 336 | Qwen3-VL-4B-InstructOpen weightsQwen · Qwen3 · best of 2 rows | 7.05 | Independent | reasoningonversion4.3 | Partially comparable-0.61 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 337 | gpt-4-turboClosedOpenAI · GPT 4 | 7.04 | Independent | version4.3 | Comparable-0.62 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 338 | Command AOpen weightsCohere · Command | 6.96 | Independent | version4.3 | Comparable-0.70 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 338 | Mistral Small 3.2Open weightsMistral AI · Mistral | 6.96 | Independent | version4.3 | Comparable-0.70 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 338 | nova-proClosedAmazon Web Services · Nova | 6.96 | Independent | version4.3 | Comparable-0.70 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 341 | Qwen3.5-2BOpen weightsQwen · Qwen3.5 · best of 2 rows | 6.94 | Independent | version4.3 | Comparable-0.72 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 341 | llama-3-1-nemotron-instruct-70bOpen weightsNVIDIA · Llama 3.1 | 6.94 | Independent | version4.3 | Comparable-0.72 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 343 | Llama 3.1 8BRestricted weightsMeta AI · Llama 3.1 | 6.93 | Independent | version4.3 | Comparable-0.73 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 344 | grok-betaClosedSpaceXAI · Grok | 6.89 | Independent | version4.3 | Comparable-0.77 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 345 | qwen2.5-32b-instructOpen weightsAlibaba Group · Qwen2.5 | 6.87 | Independent | version4.3 | Comparable-0.79 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 346 | Mistral Large 2.0Open weightsMistral AI · Mistral | 6.80 | Independent | version4.3 | Comparable-0.86 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 347 | Qwen2.5 Coder 32B InstructOpen weightsQwen · Qwen2.5 | 6.74 | Independent | version4.3 | Comparable-0.92 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 348 | gpt-4ClosedOpenAI · GPT 4 | 6.70 | Independent | version4.3 | Comparable-0.96 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 348 | qwen3-14b-instructOpen weightsAlibaba Group · Qwen3 | 6.70 | Independent | version4.3 | Comparable-0.96 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 350 | Mistral Small 3Open weightsMistral AI · Mistral | 6.67 | Independent | version4.3 | Comparable-0.99 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 350 | nova-liteClosedAmazon Web Services · Nova | 6.67 | Independent | version4.3 | Comparable-0.99 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 352 | gpt-4o-miniClosedOpenAI · GPT 4 | 6.66 | Independent | version4.3 | Comparable-1.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 352 | hermes-4-llama-3-1-70bOpen weightsNous Research · Llama 3.1 | 6.66 | Independent | version4.3 | Comparable-1.00 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 354 | deepseek-v2-5Open weightsDeepSeek · DeepSeek | 6.62 | Independent | version4.3 | Comparable-1.04 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 355 | qwen3-4b-instructOpen weightsAlibaba Group · Qwen3 | 6.61 | Independent | version4.3 | Comparable-1.05 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 356 | Llama-3.1-70BOpen weightsMeta AI · Llama 3.1 | 6.60 | Independent | version4.3 | Comparable-1.06 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 357 | granite-4.1-8bOpen weightsIBM · Granite 4.1 | 6.57 | Independent | version4.3 | Comparable-1.09 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 358 | gemini-2-0-flash-thinking-exp-1219ClosedGoogle · Gemini 2.0 | 6.56 | Independent | version4.3 | Comparable-1.10 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 358 | sarvam-30bOpen weightsSarvam | 6.56 | Independent | version4.3 | Comparable-1.10 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 360 | deepseek-v2-5-sep-2024Open weightsDeepSeek · DeepSeek | 6.55 | Independent | version4.3 | Comparable-1.11 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 361 | Mistral SabaClosedMistral AI · Mistral | 6.49 | Independent | version4.3 | Comparable-1.17 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 362 | deepseek-r1-distill-llama-8bOpen weightsDeepSeek · Llama | 6.48 | Independent | version4.3 | Comparable-1.18 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 363 | olmo-3-32b-thinkOpen weightsAllen Institute for AI · OLMo 3 | 6.47 | Independent | version4.3 | Comparable-1.19 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 364 | Llama 4 ScoutOpen weightsMeta AI · Llama 4 | 6.45 | Independent | version4.3 | Comparable-1.21 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 365 | gemini-1-5-pro-may-2024ClosedGoogle · Gemini 1.5 | 6.44 | Independent | version4.3 | Comparable-1.22 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 365 | r1-1776Open weightsPerplexity AI | 6.44 | Independent | version4.3 | Comparable-1.22 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 367 | qwen-turboClosedAlibaba Group · Qwen | 6.43 | Independent | version4.3 | Comparable-1.23 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 367 | reka-flashClosedrekaai | 6.43 | Independent | version4.3 | Comparable-1.23 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 369 | llama-3-2-instruct-90b-visionOpen weightsMeta AI · Llama 3.2 | 6.41 | Independent | version4.3 | Comparable-1.25 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 369 | solar-miniOpen weightsUpstage · Solar | 6.41 | Independent | version4.3 | Comparable-1.25 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 371 | Qwen3 14BOpen weightsQwen · Qwen3 | 6.39 | Independent | version4.3 | Comparable-1.27 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 372 | celeris-1ClosedCeleris | 6.35 | Independent | version4.3 | Comparable-1.31 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 373 | grok-1Open weightsxAI · Grok 1 | 6.34 | Independent | version4.3 | Comparable-1.32 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 374 | phi-4-miniOpen weightsMicrosoft · Phi4 | 6.33 | Independent | version4.3 | Comparable-1.33 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 374 | qwen2-72b-instructOpen weightsAlibaba Group · Qwen2 | 6.33 | Independent | version4.3 | Comparable-1.33 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 376 | gemini-1-5-flash-8bClosedGoogle · Gemini 1.5 | 6.15 | Independent | version4.3 | Comparable-1.51 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 377 | Qwen3.5-0.8BOpen weightsQwen · Qwen3.5 · best of 2 rows | 6.13 | Independent | version4.3 | Comparable-1.53 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 378 | deephermes-3-mistral-24b-previewOpen weightsNous Research · Mistral | 6.07 | Independent | version4.3 | Comparable-1.59 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 378 | jamba-1-7-largeOpen weightsAI21 Labs · Jamba 1.7 | 6.07 | Independent | version4.3 | Comparable-1.59 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 380 | granite-4-0-h-smallOpen weightsIBM · Granite 4.0 | 6.05 | Independent | version4.3 | Comparable-1.61 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 381 | Ministral 3 14BOpen weightsMistral AI · Ministral 3 | 6.04 | Independent | version4.3 | Comparable-1.62 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 382 | jamba-1-5-largeOpen weightsAI21 Labs · Jamba 1.5 | 6.01 | Independent | version4.3 | Comparable-1.65 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 383 | Hermes 3 70B InstructOpen weightsNous Research · Hermes 3 | 6 | Independent | version4.3 | Comparable-1.66 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 384 | qwen3-8b-instructOpen weightsAlibaba Group · Qwen3 | 5.99 | Independent | version4.3 | Comparable-1.67 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 385 | deepseek-coder-v2Open weightsDeepSeek · DeepSeek | 5.98 | Independent | version4.3 | Comparable-1.68 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 386 | jamba-1-6-largeOpen weightsAI21 Labs · Jamba 1.6 | 5.97 | Independent | version4.3 | Comparable-1.69 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 386 | olmo-2-32bOpen weightsAllen Institute for AI · OLMo 2 | 5.97 | Independent | version4.3 | Comparable-1.69 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 388 | lfm2-24b-a2bOpen weightsLiquid AI · LFM2 | 5.95 | Independent | version4.3 | Comparable-1.71 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 389 | gemini-1-5-flash-may-2024ClosedGoogle · Gemini 1.5 | 5.94 | Independent | version4.3 | Comparable-1.72 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 390 | Phi 4Open weightsMicrosoft · Phi4 | 5.92 | Independent | version4.3 | Comparable-1.74 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 391 | claude-3-sonnetClosedAnthropic · Claude 3 | 5.88 | Independent | version4.3 | Comparable-1.78 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 391 | nova-microClosedAmazon Web Services · Nova | 5.88 | Independent | version4.3 | Comparable-1.78 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 393 | granite-4.1-3bOpen weightsIBM · Granite 4.1 | 5.86 | Independent | version4.3 | Comparable-1.80 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 394 | mistral-smallOpen weightsMistral AI · Mistral | 5.85 | Independent | version4.3 | Comparable-1.81 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 395 | gemini-1-0-ultraClosedGoogle · Gemini 1.0 | 5.84 | Independent | version4.3 | Comparable-1.82 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 396 | phi-3-miniOpen weightsMicrosoft · Phi3 | 5.82 | Independent | version4.3 | Comparable-1.84 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 397 | gemma-3n-e4b-preview-0520Open weightsGoogle · Gemma 3 | 5.81 | Independent | version4.3 | Comparable-1.85 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 397 | phi-4-multimodalOpen weightsMicrosoft · Phi4 | 5.81 | Independent | version4.3 | Comparable-1.85 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 399 | Qwen2.5-Coder-7BOpen weightsQwen · Qwen2.5 | 5.79 | Independent | version4.3 | Comparable-1.87 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 400 | Mistral LargeClosedMistral AI · Mistral | 5.76 | Independent | version4.3 | Comparable-1.90 | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
One row per canonical model — its best current row inside this comparability group (effort variants are folded into the model). Bars are relative to the page's best score. “vs leader” reads comparability: partially comparable = same task, conditions differ (reasoning effort, temperature, judge). Rules →