Updated 2 h ago · first seen 11 Sept 2026
model_01M29A6S4J0SEWPFC7Q4G4AQF3
Specification
No structured attributes yet.
Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →
Provenance
Attributed facts
0
Source tiers
—
Freshest observation
—
Conflicts
None
Modalities
Modalities unavailable.
Capabilities
Tool calling
Unavailable
Structured output
Unavailable
Reasoning
Unavailable
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
No capability flags have been observed from a source yet — we do not infer them.
No structured attributes yet.
| Benchmark | Score | Metric | Config | Evaluated | Source |
|---|---|---|---|---|---|
| LiveBench | 76.22% | global_average | release=2026-06-25 · aggregation=mean of category averages; category = mean of its subtasks · livebench_model_id=claude-opus-4-8-max-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 89.19% | category:Reasoning | release=2026-06-25 · subtasks=["theory_of_mind","zebra_puzzle","spatial","logic_with_navigation"] · livebench_model_id=claude-opus-4-8-max-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 81.83% | category:Coding | release=2026-06-25 · subtasks=["code_generation","code_completion"] · livebench_model_id=claude-opus-4-8-max-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 50.51% | category:Agentic Coding | release=2026-06-25 · subtasks=["javascript","typescript","python"] · livebench_model_id=claude-opus-4-8-max-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 94.32% | category:Mathematics | release=2026-06-25 · subtasks=["AMPS_Hard","integrals_with_game","math_comp","olympiad"] · livebench_model_id=claude-opus-4-8-max-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 66.03% | category:Data Analysis | release=2026-06-25 · subtasks=["consecutive_events","tablejoin","tablereformat"] · livebench_model_id=claude-opus-4-8-max-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 79.66% | category:Language | release=2026-06-25 · subtasks=["connections","plot_unscrambling","typos"] · livebench_model_id=claude-opus-4-8-max-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 72.03% | category:IF | release=2026-06-25 · subtasks=["paraphrase","simplify","story_generation","summarize"] · livebench_model_id=claude-opus-4-8-max-effort | 25 Jun 2026 | livebench.aiT2 |
Scores are reported as published, with their evaluation configuration (harness, prompting, judge). Results with different configs are not directly comparable — see methodology.
Current prices
No current prices recorded
Price history
No hardware estimate available
No lineage recorded
Papers 0
No papers linked yet.
Repositories 0
No repositories linked yet.
As of
Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.
Claim history
No claims recorded yet
New model: Claude 4.8 Opus Thinking Max Effort (Anthropic)
livebench_leaderboardClaude 4.8 Opus Thinking Max Effort scores 72.034% on LiveBench
livebench_leaderboardClaude 4.8 Opus Thinking Max Effort scores 79.659% on LiveBench
livebench_leaderboardClaude 4.8 Opus Thinking Max Effort scores 66.033% on LiveBench
livebench_leaderboardClaude 4.8 Opus Thinking Max Effort scores 94.317% on LiveBench
livebench_leaderboardClaude 4.8 Opus Thinking Max Effort scores 50.505% on LiveBench
livebench_leaderboardClaude 4.8 Opus Thinking Max Effort scores 81.828% on LiveBench
livebench_leaderboardClaude 4.8 Opus Thinking Max Effort scores 89.192% on LiveBench
livebench_leaderboardClaude 4.8 Opus Thinking Max Effort scores 76.224% on LiveBench
livebench_leaderboard
| Source | Document | Type | Tier | Last observed | Snapshots |
|---|---|---|---|---|---|
| LiveBench | livebench.ai/table_2026_06_25.csv | leaderboard | T2· Quality secondary | 2 h ago | 1 |
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.