Updated 2 h ago · first seen 11 Sept 2026
model_01M29A6S1SN9GJRTM9M8GPS7QX
Specification
No structured attributes yet.
Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →
Provenance
Attributed facts
0
Source tiers
—
Freshest observation
—
Conflicts
None
Modalities
Modalities unavailable.
Capabilities
Tool calling
Unavailable
Structured output
Unavailable
Reasoning
Unavailable
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
No capability flags have been observed from a source yet — we do not infer them.
No structured attributes yet.
| Benchmark | Score | Metric | Config | Evaluated | Source |
|---|---|---|---|---|---|
| LiveBench | 74.52% | global_average | release=2026-06-25 · aggregation=mean of category averages; category = mean of its subtasks · livebench_model_id=claude-opus-4-6-thinking-auto-high-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 88.67% | category:Reasoning | release=2026-06-25 · subtasks=["theory_of_mind","zebra_puzzle","spatial","logic_with_navigation"] · livebench_model_id=claude-opus-4-6-thinking-auto-high-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 78.18% | category:Coding | release=2026-06-25 · subtasks=["code_generation","code_completion"] · livebench_model_id=claude-opus-4-6-thinking-auto-high-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 48.99% | category:Agentic Coding | release=2026-06-25 · subtasks=["javascript","typescript","python"] · livebench_model_id=claude-opus-4-6-thinking-auto-high-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 89.32% | category:Mathematics | release=2026-06-25 · subtasks=["AMPS_Hard","integrals_with_game","math_comp","olympiad"] · livebench_model_id=claude-opus-4-6-thinking-auto-high-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 69.89% | category:Data Analysis | release=2026-06-25 · subtasks=["consecutive_events","tablejoin","tablereformat"] · livebench_model_id=claude-opus-4-6-thinking-auto-high-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 83.27% | category:Language | release=2026-06-25 · subtasks=["connections","plot_unscrambling","typos"] · livebench_model_id=claude-opus-4-6-thinking-auto-high-effort | 25 Jun 2026 | livebench.aiT2 |
| LiveBench | 63.31% | category:IF | release=2026-06-25 · subtasks=["paraphrase","simplify","story_generation","summarize"] · livebench_model_id=claude-opus-4-6-thinking-auto-high-effort | 25 Jun 2026 | livebench.aiT2 |
Scores are reported as published, with their evaluation configuration (harness, prompting, judge). Results with different configs are not directly comparable — see methodology.
Current prices
No current prices recorded
Price history
No hardware estimate available
No lineage recorded
Papers 0
No papers linked yet.
Repositories 0
No repositories linked yet.
As of
Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.
Claim history
No claims recorded yet
New model: Claude 4.6 Opus Thinking High Effort (Anthropic)
livebench_leaderboardClaude 4.6 Opus Thinking High Effort scores 63.312% on LiveBench
livebench_leaderboardClaude 4.6 Opus Thinking High Effort scores 83.27% on LiveBench
livebench_leaderboardClaude 4.6 Opus Thinking High Effort scores 69.893% on LiveBench
livebench_leaderboardClaude 4.6 Opus Thinking High Effort scores 89.317% on LiveBench
livebench_leaderboardClaude 4.6 Opus Thinking High Effort scores 48.99% on LiveBench
livebench_leaderboardClaude 4.6 Opus Thinking High Effort scores 78.184% on LiveBench
livebench_leaderboardClaude 4.6 Opus Thinking High Effort scores 88.673% on LiveBench
livebench_leaderboardClaude 4.6 Opus Thinking High Effort scores 74.52% on LiveBench
livebench_leaderboard
| Source | Document | Type | Tier | Last observed | Snapshots |
|---|---|---|---|---|---|
| LiveBench | livebench.ai/table_2026_06_25.csv | leaderboard | T2· Quality secondary | 2 h ago | 1 |
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.