Graph explorer Model
Benchmark graph around gpt-5.6-luna
Benchmarks and the models evaluated on them. Click a node to inspect it, double-click to expand, drag to pan, wheel to zoom.
15 nodes · 14 edges
15 nodes · 14 edges
100% · drag to pan · wheel or pinch to zoom · double-click a node to expand
Every node, as a list (15)
- Modelgpt-5.6-lunaOpenAI
- BenchmarkLiveBench Reasoning
- BenchmarkLiveBench Language
- BenchmarkArtificial Analysis Intelligence Index
- BenchmarkMMMU-Pro
- BenchmarkHumanity's Last Exam
- BenchmarkLiveBench
- BenchmarkGPQA Diamond
- BenchmarkLiveBench Data Analysis
- BenchmarkLiveBench Mathematics
- BenchmarkTerminal-Bench
- BenchmarkLiveBench Coding
- BenchmarkSciCode
- BenchmarkLiveBench Agentic Coding
- BenchmarkLiveBench Instruction Following
Direct relations of gpt-5.6-luna, as a list
- Evaluated on 14
- BenchmarkTerminal-BenchBenchmarkArtificial Analysis Intelligence IndexBenchmarkGPQA DiamondBenchmarkHumanity's Last ExamBenchmarkSciCodeBenchmarkMMMU-ProBenchmarkLiveBench LanguageBenchmarkLiveBench Instruction FollowingBenchmarkLiveBench ReasoningBenchmarkLiveBench CodingBenchmarkLiveBench Agentic CodingBenchmarkLiveBench MathematicsBenchmarkLiveBench Data AnalysisBenchmarkLiveBench