Skip to content
AI Atlas
BenchmarkActivecategory · compositefamily · artificial-analysis-intelligence-index · variant index

Artificial Analysis Intelligence Index

artificialanalysis.ai

composite of several evaluations run by Artificial Analysis

quality57

Updated 27 min ago · first seen 11 Sept 2026

Metric
index
Current results
638
Models
477
Current leader
Claude Fable 5.1 53.4

Score history · Gemini 3.5 Flash 3 rows

Not enough history to chart — 3 observations, all dated 11 Sept 2026. Rows under different configurations count separately; the list below shows each one.

  • 32.98aa_slug=gemini-3-5-flash · version=4.3 · estimated=false11 Sept 2026
  • 23.85aa_slug=gemini-3-5-flash-minimal · version=4.3 · estimated=true · aa_variant_slug=gemini-3-5-flash-minimal11 Sept 2026
  • 33.63aa_slug=gemini-3-5-flash-medium · version=4.3 · estimated=true · aa_variant_slug=gemini-3-5-flash-medium11 Sept 2026

Back to the leaderboard

Frontier over time · index

8 leader changes recorded, all dated 11 Sept 2026 — the frontier line needs at least two distinct dates. The corpus is young: every result was first observed on the same day, so leader changes will separate in time as sources are re-crawled.

  1. 53.4Claude Fable 5.1 Anthropic Independent11 Sept 2026
  2. 53.2Claude Fable 5.1 Anthropic Independent11 Sept 2026
  3. 51.2Claude Fable 5.1 Anthropic Independent11 Sept 2026
  4. 51.0gpt-6-astra OpenAI Independent11 Sept 2026
  5. 49.7gpt-6-astra OpenAI Independent11 Sept 2026
  6. 45.1Claude Opus 5 Anthropic Independent11 Sept 2026
  7. 14.6grok-3-mini-reasoning SpaceXAI Independent11 Sept 2026
  8. 5.71llama-2-chat-7b Meta AI Independent11 Sept 2026

Includes closed rows (history). A point is emitted whenever a result beats every earlier result of the same group, ordered by evaluated_at when the source publishes it, else observed_at.

Leaderboard 477 models

Select models with +, then Compare.

No result in this group with these filters

Relax the trust / organization filters or pick another comparability group.