Artificial Analysis Intelligence Index
composite of several evaluations run by Artificial Analysis
Updated 9 h ago · first seen 11 Sept 2026
bench_01M293SPG0EAFKRT01VFHFK9SC
- Metric
- index
- Direction
- Higher is better
- Results
- 638
- Leader
- Claude Fable 5.1 53.37
Score history · gpt-5-5-non-reasoning 1 row
Not enough history to chart — a single observation (23.17 on 11 Sept 2026). Rows under different configurations count separately; the list below shows each one.
- 23.17aa_slug=gpt-5-5-non-reasoning · version=4.3 · estimated=true11 Sept 2026
Leaderboard 638 current results
Select models with +, then open Compare.
| # | Model | Score | Config | Evaluated | Source | Actions |
|---|---|---|---|---|---|---|
| #301grok-3SpaceXAI | 12.11 | aa_slug=grok-3 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #302Seed-OSS-36B-InstructByteDance | 12.1 | aa_slug=seed-oss-36b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #303qwen3-235b-a22b-instruct-2507Alibaba Group | 12 | aa_slug=qwen3-235b-a22b-instruct-2507 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #304qwen3-coder-480b-a35b-instructAlibaba Group | 11.9 | aa_slug=qwen3-coder-480b-a35b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #305qwen3-vl-32b-reasoningAlibaba Group | 11.87 | aa_slug=qwen3-vl-32b-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #306ling-3-0-tinyinclusionAI | 11.87 | aa_slug=ling-3-0-tiny · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #307Magistral Medium 1.2Mistral AI | 11.83 | aa_slug=magistral-medium-2509 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #308Sonar Reasoning ProPerplexity AI | 11.82 | aa_slug=sonar-reasoning-pro · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #309Granite 4.2 8BIBM | 11.81 | aa_slug=granite-4-2-8b · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #310nova-2-0-lite-reasoning-lowAmazon Web Services | 11.8 | aa_slug=nova-2-0-lite-reasoning-low · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #311hypernova-60bMultiverse Computing | 11.74 | aa_slug=hypernova-60b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #312MiniMax-M1-80kMiniMax | 11.73 | aa_slug=minimax-m1-80k · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #313gpt-5-4-nano-non-reasoningOpenAI | 11.69 | aa_slug=gpt-5-4-nano-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #314nemotron-cascade-2-30b-a3bNVIDIA | 11.67 | aa_slug=nemotron-cascade-2-30b-a3b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #315gemini-2-5-flash-reasoning-04-2025Google | 11.65 | aa_slug=gemini-2-5-flash-reasoning-04-2025 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #316Mercury 2Inception | 11.51 | aa_slug=mercury-2 · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #317k2-think-v2MBZUAI Institute of Foundation Models | 11.5 | aa_slug=k2-think-v2 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #318longcat-flash-liteLongCat | 11.47 | aa_slug=longcat-flash-lite · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #319gpt-5-minimalOpenAI | 11.45 | aa_slug=gpt-5-minimal · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #320Mistral Small 4Mistral AI | 11.45 | aa_slug=mistral-small-4 · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #321deepseek-r1-0120DeepSeek | 11.41 | aa_slug=deepseek-r1-0120 · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #322o1 PreviewOpenAI | 11.38 | aa_slug=o1-preview · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #323hyperclova-x-seed-think-32bNaver | 11.36 | aa_slug=hyperclova-x-seed-think-32b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #324grok-4-1-fastSpaceXAI | 11.28 | aa_slug=grok-4-1-fast · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #325GLM 4.6VZ.ai (Zhipu AI) | 11.22 | aa_slug=glm-4-6v-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #326k-exaone-non-reasoningLG AI Research | 11.21 | aa_slug=k-exaone-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #327qwen3-next-80b-a3b-reasoningAlibaba Group | 11.2 | aa_slug=qwen3-next-80b-a3b-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #328gpt-5-4-mini-non-reasoningOpenAI | 11.14 | aa_slug=gpt-5-4-mini-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #329nova-2-0-omni-reasoning-lowAmazon Web Services | 11.11 | aa_slug=nova-2-0-omni-reasoning-low · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #330GLM 4.5 AirZ.ai (Zhipu AI) | 11.09 | aa_slug=glm-4-5-air · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #331grok-4-fastSpaceXAI | 11.06 | aa_slug=grok-4-fast · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #332mi-dm-k-2-5-pro-dec28Korea Telecom | 11.04 | aa_slug=mi-dm-k-2-5-pro-dec28 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #333o3 Mini HighOpenAI | 10.96 | aa_slug=o3-mini-high · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #334ring-1tinclusionAI | 10.9 | aa_slug=ring-1t · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #335Trinity Large ThinkingArcee AI | 10.88 | aa_slug=trinity-large-thinking · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #336g9v3-3bAI9Stars | 10.84 | aa_slug=g9v3-3b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #337qwen3-5-4b-non-reasoningAlibaba Group | 10.81 | aa_slug=qwen3-5-4b-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #338intellect-3Prime Intellect | 10.61 | aa_slug=intellect-3 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #339glm-4-7-flash-non-reasoningZ.ai (Zhipu AI) | 10.56 | aa_slug=glm-4-7-flash-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #340gpt-5-chatgptOpenAI | 10.44 | aa_slug=gpt-5-chatgpt · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #341solar-open-100b-reasoningUpstage | 10.36 | aa_slug=solar-open-100b-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #342grok-3-reasoningSpaceXAI | 10.36 | aa_slug=grok-3-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #343gemini-2-5-flash-lite-preview-09-2025-reasoningGoogle | 10.35 | aa_slug=gemini-2-5-flash-lite-preview-09-2025-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #344nemotron-3-nano-omni-30b-a3bNVIDIA | 10.25 | aa_slug=nemotron-3-nano-omni-30b-a3b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #345gpt-oss-120b-lowOpenAI | 10.21 | aa_slug=gpt-oss-120b-low · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #346gpt-4.1-miniOpenAI | 10.16 | aa_slug=gpt-4-1-mini · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #347Qwen3 Coder NextQwen | 10.05 | aa_slug=qwen3-coder-next · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #348MiniMax-M1-40kMiniMax | 9.99 | aa_slug=minimax-m1-40k · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #349nova-2-0-proAmazon Web Services | 9.96 | aa_slug=nova-2-0-pro · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #350gpt-oss-20b-lowOpenAI | 9.95 | aa_slug=gpt-oss-20b-low · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #351Qwen3 VL 235B A22B InstructQwen | 9.94 | aa_slug=qwen3-vl-235b-a22b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #352gpt-5-mini-minimalOpenAI | 9.91 | aa_slug=gpt-5-mini-minimal · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #353Mistral Medium 3.1Mistral AI | 9.89 | aa_slug=mistral-medium-3-1 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #354k2-v2MBZUAI Institute of Foundation Models | 9.87 | aa_slug=k2-v2 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #355Gemini 2.5 FlashGoogle | 9.85 | aa_slug=gemini-2-5-flash · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #356qwen3-30b-a3b-2507-reasoningAlibaba Group | 9.81 | aa_slug=qwen3-30b-a3b-2507-reasoning · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #357o1-miniOpenAI | 9.77 | aa_slug=o1-mini · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #358DeepSeek V3 0324DeepSeek | 9.72 | aa_slug=deepseek-v3-0324 · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #359Mistral Large 3Mistral AI | 9.71 | aa_slug=mistral-large-3 · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #360ling-2-6-flashinclusionAI | 9.68 | aa_slug=ling-2-6-flash · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #361Qwen3 Next 80B A3B InstructQwen | 9.64 | aa_slug=qwen3-next-80b-a3b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #362Qwen3 Coder 30B A3B InstructQwen | 9.59 | aa_slug=qwen3-coder-30b-a3b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #363tri-21b-think-previewTrillion Labs | 9.59 | aa_slug=tri-21b-think-preview · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #364GPT-4.5 PreviewOpenAI | 9.58 | aa_slug=gpt-4-5 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #365diffusiongemma-26b-a4bGoogle | 9.51 | aa_slug=diffusiongemma-26b-a4b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #366qwen3-235b-a22b-instruct-reasoningAlibaba Group | 9.5 | aa_slug=qwen3-235b-a22b-instruct-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #367QwQ-32BAlibaba Group | 9.47 | aa_slug=qwq-32b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #368qwen3-vl-30b-a3b-reasoningAlibaba Group | 9.45 | aa_slug=qwen3-vl-30b-a3b-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #369gemini-2.0-flash-thinking-exp-01-21Google | 9.42 | aa_slug=gemini-2-0-flash-thinking-exp-0121 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #370Devstral 2Mistral AI | 9.41 | aa_slug=devstral-2 · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #371gemma-4-12b-non-reasoningGoogle | 9.38 | aa_slug=gemma-4-12b-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #372gemini-2-5-flash-lite-preview-09-2025Google | 9.34 | aa_slug=gemini-2-5-flash-lite-preview-09-2025 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #373Llama 4 MaverickMeta AI | 9.3 | aa_slug=llama-4-maverick · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #374motif-2-12-7bMotif Technologies | 9.19 | aa_slug=motif-2-12-7b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #375ling-1tinclusionAI | 9.17 | aa_slug=ling-1t · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #376nova-premierAmazon Web Services | 9.16 | aa_slug=nova-premier · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #377solar-pro-2-preview-reasoningUpstage | 9.07 | aa_slug=solar-pro-2-preview-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #378granite-4.2-3bIBM | 9.06 | aa_slug=granite-4-2-3b · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #379Mistral Medium 3Mistral AI | 9.05 | aa_slug=mistral-medium-3 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #380magistral-mediumMistral AI | 9.05 | aa_slug=magistral-medium · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #381gpt-oss-20bOpenAI | 9.04 | aa_slug=gpt-oss-20b · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #382k2-v2-mediumMBZUAI Institute of Foundation Models | 9.02 | aa_slug=k2-v2-medium · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #383devstral-mediumMistral AI | 9.01 | aa_slug=devstral-medium · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #384llama-nemotron-super-49b-v1-5-reasoningNVIDIA | 9.01 | aa_slug=llama-nemotron-super-49b-v1-5-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #385mistral-small-4-non-reasoningMistral AI | 8.99 | aa_slug=mistral-small-4-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #386tri-21b-think-v0-5Trillion Labs | 8.99 | aa_slug=tri-21b-think-v0-5 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #387gpt-4o-chatgpt-03-25OpenAI | 8.96 | aa_slug=gpt-4o-chatgpt-03-25 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #388Gemini 2.0 FlashGoogle | 8.94 | aa_slug=gemini-2-0-flash · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #389Claude Haiku 3.5Anthropic | 8.94 | aa_slug=claude-3-5-haiku · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #390llama-3-3-nemotron-super-49b-reasoningNVIDIA | 8.93 | aa_slug=llama-3-3-nemotron-super-49b-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #391gemma-4-E4BGoogle | 8.91 | aa_slug=gemma-4-e4b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #392nvidia-nemotron-3-nano-30b-a3b-reasoningNVIDIA | 8.9 | aa_slug=nvidia-nemotron-3-nano-30b-a3b-reasoning · version=4.3 · estimated=false | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #393minicpm5-1bOpenBMB | 8.8 | aa_slug=minicpm5-1b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #394qwen3-4b-2507-instruct-reasoningAlibaba Group | 8.8 | aa_slug=qwen3-4b-2507-instruct-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #395sarvam-105bSarvam | 8.79 | aa_slug=sarvam-105b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #396gemini-2-0-pro-experimental-02-05Google | 8.75 | aa_slug=gemini-2-0-pro-experimental-02-05 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #397nova-2-0-liteAmazon Web Services | 8.74 | aa_slug=nova-2-0-lite · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #398Devstral Small 1.0Mistral AI | 8.74 | aa_slug=devstral-small-2505 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #399claude-3-opusAnthropic | 8.72 | aa_slug=claude-3-opus · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #400minicpm5-1b-non-reasoningOpenBMB | 8.69 | aa_slug=minicpm5-1b-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History |
Scores are reported as published, with their evaluation configuration (harness, prompting, judge). The bar is relative to the best score on this page. Results with different configs are not directly comparable — see methodology.
The config filter matches a value inside each result's configuration (server-side, `config=` on the API). Chips are the values shared by several rows on the first page; per-model identifiers are not offered.
Definition
- Category
- composite
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
- Task
- composite of several evaluations run by Artificial Analysis
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
- Metric
- index
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
- Direction
- Higher is better
- Known limitations
- Proprietary composite; methodology versions change.
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
- Website
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
Each value shows its source, tier and observation time. Missing rows mean no source stated them. How results are recorded →
- Evaluated models
- llama-2-chat-7b, grok-3-mini-reasoning, claude-opus-5-medium, gemma-4-12B, deepseek-v4-flash, grok-4-3-low, llama-65b, gemini-2.0-flash-exp +630(638 total)
As of
Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.
Claim history
Categorycategory1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| composite | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Known limitationsknown_limitations1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| Proprietary composite; methodology versions change. | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Metricmetric1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| index | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Tasktask1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| composite of several evaluations run by Artificial Analysis | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Websitewebsite1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| https://artificialanalysis.ai | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
New benchmark: Artificial Analysis Intelligence Index
registry
| Source | Document | Type | Tier | Last observed | Snapshots |
|---|---|---|---|---|---|
| Artificial Analysis | artificialanalysis.ai/leaderboards/models | leaderboard | T2· Quality secondary | 8 h ago | 1 |
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.