Updated 13 min ago · first seen 11 Sept 2026
- Metric
- accuracy · % ↑
- Current results
- 821
- Models
- 374
- Current leader
- gpt-5.6-sol 65.9%
Score history · MiniMax-M1-80k 1 row
Not enough history to chart — a single observation (3.03% on 11 Sept 2026). Rows under different configurations count separately; the list below shows each one.
- 3.03%aa_slug=minimax-m1-80k · variant=hard · evaluator=Artificial Analysis · index_version=4.311 Sept 2026
Frontier over time · accuracy · variant=hard · evaluator=Artificial Analysis
10 leader changes recorded, all dated 11 Sept 2026 — the frontier line needs at least two distinct dates. The corpus is young: every result was first observed on the same day, so leader changes will separate in time as sources are re-crawled.
- 65.9%gpt-5.6-sol OpenAI Independent11 Sept 2026
- 61.4%gpt-5.6-sol OpenAI Independent11 Sept 2026
- 53.0%Claude Sonnet 4.6 Anthropic Independent11 Sept 2026
- 51.5%Claude Opus 4.7 Anthropic Independent11 Sept 2026
- 50.8%Z.ai GLM 5.2 Z.ai (Zhipu AI) Independent11 Sept 2026
- 49.2%KAT-Coder-Pro V2 Kwaipilot Independent11 Sept 2026
- 33.3%GPT-5.1-Codex Mini OpenAI Independent11 Sept 2026
- 26.5%Grok 4.3 xAI Independent11 Sept 2026
Includes closed rows (history). A point is emitted whenever a result beats every earlier result of the same group, ordered by evaluated_at when the source publishes it, else observed_at.
Leaderboard 315 models
Select models with +, then Compare.
| # | Model | Score | Trust | Configuration | vs leader | Evaluated | Source | Actions |
|---|---|---|---|---|---|---|---|---|
| 99 | command-a-plusOpen weightsCohere · Command | 25% | Independent | group defaults | Comparable0.00 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 99 | deepseek-v3-2-0925Open weightsDeepSeek · DeepSeek | 25% | Independent | group defaults | Comparable0.00 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 99 | ernie-5-0-thinking-previewClosedBaidu · ERNIE 5.0 | 25% | Independent | group defaults | Comparable0.00 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 104 | Gemini 3.1 Flash-Lite PreviewClosedGoogle · Gemini | 24.2% | Independent | group defaults | Comparable-0.76 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 104 | Qwen3 MaxClosedQwen · Qwen3 · best of 2 rows | 24.2% | Independent | reasoningon | Partially comparable-0.76 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 104 | Qwen3.5-9BOpen weightsQwen · Qwen3.5 · best of 2 rows | 24.2% | Independent | group defaults | Comparable-0.76 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 104 | grok-4-1-fastClosedSpaceXAI · Grok 4.1 · best of 2 rows | 24.2% | Independent | reasoningon | Partially comparable-0.76 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 104 | nova-2-0-proClosedAmazon Web Services · Nova 2.0 · best of 3 rows | 24.2% | Independent | reasoningonreasoning_effortmedium | Partially comparable-0.76 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 109 | Kimi K2 0905Open weightsMoonshot AI · Kimi | 23.5% | Independent | group defaults | Comparable-1.52 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 109 | gpt-oss-120bOpen weightsOpenAI · gpt-oss · best of 2 rows | 23.5% | Independent | group defaults | Comparable-1.52 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 109 | hypernova-60bOpen weightsMultiverse Computing | 23.5% | Independent | group defaults | Comparable-1.52 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 112 | Trinity Large ThinkingOpen weightsArcee AI | 22.7% | Independent | group defaults | Comparable-2.27 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 112 | k-exaoneOpen weightsLG AI Research · EXAONE · best of 2 rows | 22.7% | Independent | group defaults | Comparable-2.27 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 114 | GLM 4.5Open weightsZ.ai (Zhipu AI) · GLM4.5 | 22.0% | Independent | group defaults | Comparable-3.03 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 114 | GLM 4.7 FlashOpen weightsZ.ai (Zhipu AI) · GLM4.7 · best of 2 rows | 22.0% | Independent | group defaults | Comparable-3.03 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 116 | Claude 3.7 SonnetClosedAnthropic · Claude · best of 2 rows | 21.2% | Independent | group defaults | Comparable-3.79 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 116 | ling-2-6-flashOpen weightsinclusionAI | 21.2% | Independent | group defaults | Comparable-3.79 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 116 | nemotron-cascade-2-30b-a3bOpen weightsNVIDIA · Nemotron | 21.2% | Independent | group defaults | Comparable-3.79 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 116 | qwen3-5-omni-plusClosedAlibaba Group · Qwen3.5 | 21.2% | Independent | group defaults | Comparable-3.79 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 120 | GLM 4.5 AirOpen weightsZ.ai (Zhipu AI) · GLM4.5 | 20.4% | Independent | group defaults | Comparable-4.55 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 120 | exaone-4-5-33bOpen weightsLG AI Research · EXAONE 4.5 | 20.4% | Independent | group defaults | Comparable-4.55 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 122 | qwen3-max-previewClosedAlibaba Group · Qwen3 | 19.7% | Independent | group defaults | Comparable-5.30 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 123 | Devstral 2Open weightsMistral AI · Devstral 2 | 18.9% | Independent | group defaults | Comparable-6.06 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 123 | grok-4-fastClosedSpaceXAI · Grok 4 · best of 2 rows | 18.9% | Independent | reasoningon | Partially comparable-6.06 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 123 | qwen3-coder-480b-a35b-instructOpen weightsAlibaba Group · Qwen3-Coder | 18.9% | Independent | group defaults | Comparable-6.06 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 126 | Qwen3 Coder NextOpen weightsQwen · Qwen3 | 18.2% | Independent | group defaults | Comparable-6.82 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 126 | Qwen3.5-4BOpen weightsQwen · Qwen3.5 · best of 2 rows | 18.2% | Independent | group defaults | Comparable-6.82 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 126 | gemma-4-12BOpen weightsGoogle · Gemma 4 · best of 2 rows | 18.2% | Independent | group defaults | Comparable-6.82 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 126 | jt-miniClosedChina Mobile | 18.2% | Independent | group defaults | Comparable-6.82 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 130 | Grok Build 0.1ClosedxAI · Grok | 17.4% | Independent | group defaults | Comparable-7.58 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 130 | Mistral Small 4Open weightsMistral AI · Mistral · best of 2 rows | 17.4% | Independent | group defaults | Comparable-7.58 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 130 | gpt-5-nanoClosedOpenAI · GPT 5 · best of 3 rows | 17.4% | Independent | reasoning_effortmedium | Partially comparable-7.58 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 130 | grok-3-mini-reasoningClosedSpaceXAI · Grok 3 | 17.4% | Independent | group defaults | Comparable-7.58 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 130 | nova-2-0-liteClosedAmazon Web Services · Nova 2.0 · best of 4 rows | 17.4% | Independent | reasoningonreasoning_effortmedium | Partially comparable-7.58 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 130 | qwen3-max-thinking-previewClosedAlibaba Group · Qwen3 | 17.4% | Independent | group defaults | Comparable-7.58 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 136 | Devstral Small 2Open weightsMistral AI · Devstral | 16.7% | Independent | group defaults | Comparable-8.33 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 136 | cogito-v2-1-reasoningOpen weightsDeep Cogito · Cogito | 16.7% | Independent | group defaults | Comparable-8.33 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 136 | gemini-2-5-flash-preview-09-2025ClosedGoogle · Gemini 2.5 · best of 2 rows | 16.7% | Independent | reasoningon | Partially comparable-8.33 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 136 | grok-4.20-0309-non-reasoningClosedxAI · Grok | 16.7% | Independent | group defaults | Comparable-8.33 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 140 | Kimi K2 0711Open weightsMoonshot AI · Kimi | 15.9% | Independent | group defaults | Comparable-9.09 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 140 | Mistral Large 3Open weightsMistral AI · Mistral | 15.9% | Independent | group defaults | Comparable-9.09 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 140 | deepseek-r1Open weightsDeepSeek · DeepSeek-R1 | 15.9% | Independent | group defaults | Comparable-9.09 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 143 | DeepSeek V3 0324Open weightsDeepSeek · DeepSeek-V3 | 15.2% | Independent | group defaults | Comparable-9.85 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 143 | Qwen3 235B A22B Instruct 2507Open weightsQwen · Qwen3 · best of 2 rows | 15.2% | Independent | group defaults | Comparable-9.85 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 143 | Qwen3 Coder 30B A3B InstructOpen weightsQwen · Qwen3 | 15.2% | Independent | group defaults | Comparable-9.85 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 143 | o4-miniClosedOpenAI · OpenAI o-series | 15.2% | Independent | group defaults | Comparable-9.85 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 147 | GLM 4.6VOpen weightsZ.ai (Zhipu AI) · GLM4.6 · best of 2 rows | 14.4% | Independent | group defaults | Comparable-10.6 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 147 | apriel-v1-6-15b-thinkerOpen weightsServiceNow | 14.4% | Independent | group defaults | Comparable-10.6 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 149 | Gemini 2.5 FlashClosedGoogle · Gemini · best of 2 rows | 13.6% | Independent | reasoningon | Partially comparable-11.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 149 | Nemotron 3 Nano 30B A3BOpen weightsNVIDIA · Nemotron 3 · best of 2 rows | 13.6% | Independent | reasoningon | Partially comparable-11.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 149 | gpt-4.1ClosedOpenAI · GPT 4.1 | 13.6% | Independent | group defaults | Comparable-11.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 152 | Magistral Medium 1.2ClosedMistral AI · Magistral | 12.9% | Independent | group defaults | Comparable-12.1 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 152 | gemini-2-5-flash-lite-preview-09-2025ClosedGoogle · Gemini 2.5 · best of 2 rows | 12.9% | Independent | reasoningon | Partially comparable-12.1 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 152 | gpt-5-chatgptClosedOpenAI · GPT 5 | 12.9% | Independent | group defaults | Comparable-12.1 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 152 | o1ClosedOpenAI · OpenAI o-series | 12.9% | Independent | group defaults | Comparable-12.1 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 156 | hyperclova-x-seed-think-32bOpen weightsNaver · Seed | 12.1% | Independent | group defaults | Comparable-12.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 157 | Hermes 4 405BOpen weightsNous Research · Hermes 4 | 11.4% | Independent | group defaults | Comparable-13.6 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 157 | Kimi-Linear-48B-A3B-InstructOpen weightsMoonshot AI · Kimi | 11.4% | Independent | group defaults | Comparable-13.6 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 157 | Qwen3 VL 235B A22B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 11.4% | Independent | reasoningon | Partially comparable-13.6 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 157 | grok-3ClosedSpaceXAI · Grok 3 | 11.4% | Independent | group defaults | Comparable-13.6 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 161 | Mistral Medium 3.1ClosedMistral AI · Mistral | 10.6% | Independent | group defaults | Comparable-14.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 161 | apriel-v1-5-15b-thinkerOpen weightsServiceNow | 10.6% | Independent | group defaults | Comparable-14.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 161 | gpt-oss-20bOpen weightsOpenAI · gpt-oss · best of 2 rows | 10.6% | Independent | group defaults | Comparable-14.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 161 | ling-1tOpen weightsinclusionAI | 10.6% | Independent | group defaults | Comparable-14.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 161 | ling-flash-2-0Open weightsinclusionAI | 10.6% | Independent | group defaults | Comparable-14.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 161 | longcat-flash-liteOpen weightsLongCat | 10.6% | Independent | group defaults | Comparable-14.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 167 | Qwen3 Next 80B A3B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 9.85% | Independent | reasoningon | Partially comparable-15.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 167 | hermes-4-llama-3-1-405bOpen weightsNous Research · Llama 3.1 | 9.85% | Independent | group defaults | Comparable-15.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 167 | k2-v2Open weightsMBZUAI Institute of Foundation Models · best of 3 rows | 9.85% | Independent | group defaults | Comparable-15.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 170 | devstral-mediumClosedMistral AI · Devstral | 9.09% | Independent | group defaults | Comparable-15.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 170 | intellect-3Open weightsPrime Intellect | 9.09% | Independent | group defaults | Comparable-15.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 170 | kat-coder-pro-v1ClosedKwaiKAT | 9.09% | Independent | group defaults | Comparable-15.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 170 | magistral-mediumClosedMistral AI · Magistral | 9.09% | Independent | group defaults | Comparable-15.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 174 | GPT-4o (2024-08-06)ClosedOpenAI · GPT 4 | 8.33% | Independent | group defaults | Comparable-16.7 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 174 | Qwen3 VL 32B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 8.33% | Independent | group defaults | Comparable-16.7 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 174 | gemma-4-E4BOpen weightsGoogle · Gemma 4 · best of 2 rows | 8.33% | Independent | group defaults | Comparable-16.7 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 174 | gpt-4oClosedOpenAI · GPT 4 | 8.33% | Independent | group defaults | Comparable-16.7 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 174 | nemotron-3-nano-omni-30b-a3bOpen weightsNVIDIA · Nemotron 3 | 8.33% | Independent | group defaults | Comparable-16.7 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 174 | qwen3-5-omni-flashClosedAlibaba Group · Qwen3.5 | 8.33% | Independent | group defaults | Comparable-16.7 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 180 | Mistral Small 3.1Open weightsMistral AI · Mistral | 7.58% | Independent | group defaults | Comparable-17.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 180 | Solar Pro 3ClosedUpstage · Solar | 7.58% | Independent | group defaults | Comparable-17.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 180 | gpt-4.1-miniClosedOpenAI · GPT 4.1 | 7.58% | Independent | group defaults | Comparable-17.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 180 | ring-flash-2-0Open weightsinclusionAI | 7.58% | Independent | group defaults | Comparable-17.4 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | GLM 4.5VOpen weightsZ.ai (Zhipu AI) · GLM4.5 · best of 2 rows | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | Llama 4 MaverickOpen weightsMeta AI · Llama 4 | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | Llama-3.1-405BRestricted weightsMeta AI · Llama 3.1 | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | Mistral Small 3.2Open weightsMistral AI · Mistral | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | Seed-OSS-36B-InstructOpen weightsByteDance · Seed | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | k2-think-v2Open weightsMBZUAI Institute of Foundation Models | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | nova-2-0-omniClosedAmazon Web Services · Nova 2.0 · best of 3 rows | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | nova-premierClosedAmazon Web Services · Nova | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | nvidia-nemotron-3-nano-4bOpen weightsNVIDIA · Nemotron 3 | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | o3-miniClosedOpenAI · OpenAI o-series · best of 2 rows | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | qwen3-30b-a3b-instructOpen weightsAlibaba Group · Qwen3 · best of 2 rows | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 184 | ring-1tOpen weightsinclusionAI | 6.82% | Independent | group defaults | Comparable-18.2 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 196 | Devstral Small 1.0Open weightsMistral AI · Devstral | 6.06% | Independent | group defaults | Comparable-18.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 196 | Qwen3 VL 30B A3B InstructOpen weightsQwen · Qwen3 · best of 2 rows | 6.06% | Independent | group defaults | Comparable-18.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 196 | deepseek-r1-0120Open weightsDeepSeek · DeepSeek | 6.06% | Independent | group defaults | Comparable-18.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 196 | devstral-smallOpen weightsMistral AI · Devstral | 6.06% | Independent | group defaults | Comparable-18.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
| 196 | ernie-4-5-300b-a47bOpen weightsBaidu · ERNIE 4.5 | 6.06% | Independent | group defaults | Comparable-18.9 pt | obs. 11 Sept 2026 | artificialanalysis.aiT2 | History |
One row per canonical model — its best current row inside this comparability group (effort variants are folded into the model). Bars are relative to the page's best score. “vs leader” reads comparability: partially comparable = same task, conditions differ (reasoning effort, temperature, judge). Rules →