Artificial Analysis Intelligence Index
composite of several evaluations run by Artificial Analysis
Updated 9 h ago · first seen 11 Sept 2026
bench_01M293SPG0EAFKRT01VFHFK9SC
- Metric
- index
- Direction
- Higher is better
- Results
- 638 · 491 filtered
- Leader
- Claude Fable 5.1 53.37
Score history · GLM 4.7 1 row
Not enough history to chart — a single observation (22.24 on 11 Sept 2026). Rows under different configurations count separately; the list below shows each one.
- 22.24aa_slug=glm-4-7 · version=4.3 · estimated=true11 Sept 2026
Leaderboard 491 current results · config contains “true”
Select models with +, then open Compare.
| # | Model | Score | Config | Evaluated | Source | Actions |
|---|---|---|---|---|---|---|
| #401olmo-2-32bAllen Institute for AI | 5.97 | aa_slug=olmo-2-32b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #402lfm2-24b-a2bLiquid AI | 5.95 | aa_slug=lfm2-24b-a2b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #403gemini-1-5-flash-may-2024Google | 5.94 | aa_slug=gemini-1-5-flash-may-2024 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #404Phi 4Microsoft | 5.92 | aa_slug=phi-4 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #405nova-microAmazon Web Services | 5.88 | aa_slug=nova-micro · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #406claude-3-sonnetAnthropic | 5.88 | aa_slug=claude-3-sonnet · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #407granite-4.1-3bIBM | 5.86 | aa_slug=granite-4-1-3b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #408mistral-smallMistral AI | 5.85 | aa_slug=mistral-small · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #409gemini-1-0-ultraGoogle | 5.84 | aa_slug=gemini-1-0-ultra · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #410phi-3-miniMicrosoft | 5.82 | aa_slug=phi-3-mini · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #411nvidia-nemotron-nano-12b-v2-vlNVIDIA | 5.82 | aa_slug=nvidia-nemotron-nano-12b-v2-vl · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #412gemma-3n-e4b-preview-0520Google | 5.81 | aa_slug=gemma-3n-e4b-preview-0520 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #413phi-4-multimodalMicrosoft | 5.81 | aa_slug=phi-4-multimodal · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #414Qwen2.5-Coder-7BQwen | 5.79 | aa_slug=qwen2-5-coder-7b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #415Mistral LargeMistral AI | 5.76 | aa_slug=mistral-large · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #416Mixtral 8x22BMistral AI | 5.74 | aa_slug=mistral-8x22b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #417llama-2-chat-7bMeta AI | 5.71 | aa_slug=llama-2-chat-7b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #418Llama-3.2-3BMeta AI | 5.7 | aa_slug=llama-3-2-instruct-3b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #419minicpm-v4-6-1-3bOpenBMB | 5.69 | aa_slug=minicpm-v4-6-1-3b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #420jamba-reasoning-3bAI21 Labs | 5.67 | aa_slug=jamba-reasoning-3b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #421Qwen3-VL-4B-InstructQwen | 5.66 | aa_slug=qwen3-vl-4b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #422Reka Flash 3rekaai | 5.65 | aa_slug=reka-flash-3 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #423qwen1.5-110b-chatAlibaba Group | 5.65 | aa_slug=qwen1.5-110b-chat · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #424Olmo-3-7B-ThinkAllen Institute for AI | 5.62 | aa_slug=olmo-3-7b-think · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #425claude-21Anthropic | 5.59 | aa_slug=claude-21 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #426Claude 3 HaikuAnthropic | 5.58 | aa_slug=claude-3-haiku · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #427olmo-2-7bAllen Institute for AI | 5.58 | aa_slug=olmo-2-7b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #428molmo-7b-dAllen Institute for AI | 5.56 | aa_slug=molmo-7b-d · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #429ling-mini-2-0inclusionAI | 5.54 | aa_slug=ling-mini-2-0 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #430DeepSeek-R1-Distill-Qwen-1.5BDeepSeek | 5.51 | aa_slug=deepseek-r1-distill-qwen-1-5b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #431claude-2Anthropic | 5.51 | aa_slug=claude-2 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #432deepseek-v2DeepSeek | 5.51 | aa_slug=deepseek-v2 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #433Mistral Small 1.0Mistral AI | 5.5 | aa_slug=mistral-small-2402 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #434mistral-mediumMistral AI | 5.49 | aa_slug=mistral-medium · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #435GPT-3.5 TurboOpenAI | 5.49 | aa_slug=gpt-35-turbo · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #436llama-3-instruct-70bMeta AI | 5.45 | aa_slug=llama-3-instruct-70b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #437qwen-chat-72bAlibaba Group | 5.44 | aa_slug=qwen-chat-72b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #438arctic-instructSnowflake | 5.44 | aa_slug=arctic-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #439lfm-40bLiquid AI | 5.42 | aa_slug=lfm-40b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #440llama-3-2-instruct-11b-visionMeta AI | 5.41 | aa_slug=llama-3-2-instruct-11b-vision · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #441qwen3-5-0-8b-non-reasoningAlibaba Group | 5.4 | aa_slug=qwen3-5-0-8b-non-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #442palm-2Google | 5.37 | aa_slug=palm-2 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #443deepseek-coder-v2-liteDeepSeek | 5.34 | aa_slug=deepseek-coder-v2-lite · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #444gemini-1-0-proGoogle | 5.34 | aa_slug=gemini-1-0-pro · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #445sarvam-m-reasoningSarvam | 5.31 | aa_slug=sarvam-m-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #446llama-2-chat-13bMeta AI | 5.3 | aa_slug=llama-2-chat-13b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #447llama-2-chat-70bMeta AI | 5.3 | aa_slug=llama-2-chat-70b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #448deepseek-llm-67b-chatDeepSeek | 5.3 | aa_slug=deepseek-llm-67b-chat · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #449command-r-plus-04-2024Cohere | 5.3 | aa_slug=command-r-plus-04-2024 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #450openchat-35OpenChat | 5.29 | aa_slug=openchat-35 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #451dbrxDatabricks | 5.29 | aa_slug=dbrx · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #452exaone-4-0-1-2b-reasoningLG AI Research | 5.27 | aa_slug=exaone-4-0-1-2b-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #453Olmo-3-7B-InstructAllen Institute for AI | 5.24 | aa_slug=olmo-3-7b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #454exaone-4-0-1-2bLG AI Research | 5.23 | aa_slug=exaone-4-0-1-2b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #455lfm2-5-1-2b-thinkingLiquid AI | 5.22 | aa_slug=lfm2-5-1-2b-thinking · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #456jamba-1-7-miniAI21 Labs | 5.22 | aa_slug=jamba-1-7-mini · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #457jamba-1-5-miniAI21 Labs | 5.21 | aa_slug=jamba-1-5-mini · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #458LFM2.5-1.2B-InstructLiquid AI | 5.21 | aa_slug=lfm2-5-1-2b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #459lfm2-2-6bLiquid AI | 5.21 | aa_slug=lfm2-2-6b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #460qwen3-1.7b-instruct-reasoningAlibaba Group | 5.19 | aa_slug=qwen3-1.7b-instruct-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #461granite-4-0-h-nano-1bIBM | 5.19 | aa_slug=granite-4-0-h-nano-1b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #462jamba-1-6-miniAI21 Labs | 5.16 | aa_slug=jamba-1-6-mini · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #463Mixtral 8x7BMistral AI | 5.12 | aa_slug=mixtral-8x7b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #464gemma-3-270mGoogle | 5.12 | aa_slug=gemma-3-270m · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #465Granite 4.0 MicroIBM | 5.11 | aa_slug=granite-4-0-micro · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #466apertus-70b-instructSwiss AI Initiative | 5.11 | aa_slug=apertus-70b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #467deephermes-3-llama-3-1-8b-previewNous Research | 5.08 | aa_slug=deephermes-3-llama-3-1-8b-preview · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #468Mistral 7BMistral AI | 5.03 | aa_slug=mistral-7b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #469command-r-03-2024Cohere | 5.03 | aa_slug=command-r-03-2024 · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #470llama-65bMeta AI | 5.03 | aa_slug=llama-65b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #471qwen-chat-14bAlibaba Group | 5.03 | aa_slug=qwen-chat-14b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #472claude-instantAnthropic | 5.03 | aa_slug=claude-instant · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #473granite-4-0-nano-1bIBM | 5.01 | aa_slug=granite-4-0-nano-1b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #474Molmo2-8BAllen Institute for AI | 5 | aa_slug=molmo2-8b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #475lfm2-8b-a1bLiquid AI | 4.93 | aa_slug=lfm2-8b-a1b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #476granite-3-3-8b-instructIBM | 4.92 | aa_slug=granite-3-3-8b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #477qwen3-1.7b-instructAlibaba Group | 4.86 | aa_slug=qwen3-1.7b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #478Llama-3.2-1BMeta AI | 4.83 | aa_slug=llama-3-2-instruct-1b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #479apertus-8b-instructSwiss AI Initiative | 4.83 | aa_slug=apertus-8b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #480gemma-3n-e2bGoogle | 4.83 | aa_slug=gemma-3n-e2b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #481LFM2.5-VL-1.6BLiquid AI | 4.83 | aa_slug=lfm2-5-vl-1-6b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #482LFM2-1.2BLiquid AI | 4.83 | aa_slug=lfm2-1-2b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #483qwen3-0.6b-instructAlibaba Group | 4.83 | aa_slug=qwen3-0.6b-instruct · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #484granite-4-0-h-350mIBM | 4.83 | aa_slug=granite-4-0-h-350m · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #485llama-3-instruct-8bMeta AI | 4.83 | aa_slug=llama-3-instruct-8b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #486tiny-aya-globalCohere | 4.83 | aa_slug=tiny-aya-global · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #487gemma-3-1bGoogle | 4.83 | aa_slug=gemma-3-1b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #488Gemma 3 4BGoogle | 4.83 | aa_slug=gemma-3-4b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #489granite-4-0-350mIBM | 4.83 | aa_slug=granite-4-0-350m · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #490Qwen3-0.6BQwen | 4.83 | aa_slug=qwen3-0.6b-instruct-reasoning · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History | |
| #491gemma-3n-e4bGoogle | 4.83 | aa_slug=gemma-3n-e4b · version=4.3 · estimated=true | — (obs. 11 Sept 2026) | artificialanalysis.aiT2 | History |
Scores are reported as published, with their evaluation configuration (harness, prompting, judge). The bar is relative to the best score on this page. Results with different configs are not directly comparable — see methodology.
The config filter matches a value inside each result's configuration (server-side, `config=` on the API). Chips are the values shared by several rows on the first page; per-model identifiers are not offered.
Definition
- Category
- composite
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
- Task
- composite of several evaluations run by Artificial Analysis
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
- Metric
- index
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
- Direction
- Higher is better
- Known limitations
- Proprietary composite; methodology versions change.
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
- Website
Source:AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2observed 9 h agomedium
Each value shows its source, tier and observation time. Missing rows mean no source stated them. How results are recorded →
- Evaluated models
- llama-2-chat-7b, grok-3-mini-reasoning, claude-opus-5-medium, gemma-4-12B, deepseek-v4-flash, grok-4-3-low, llama-65b, gemini-2.0-flash-exp +630(638 total)
As of
Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.
Claim history
Categorycategory1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| composite | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Known limitationsknown_limitations1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| Proprietary composite; methodology versions change. | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Metricmetric1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| index | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Tasktask1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| composite of several evaluations run by Artificial Analysis | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Websitewebsite1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| https://artificialanalysis.ai | → current | current | AI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2 | medium | curated |
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
New benchmark: Artificial Analysis Intelligence Index
registry
| Source | Document | Type | Tier | Last observed | Snapshots |
|---|---|---|---|---|---|
| Artificial Analysis | artificialanalysis.ai/leaderboards/models | leaderboard | T2· Quality secondary | 7 h ago | 1 |
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.