Find a model
Describe the need; the atlas applies deterministic rules to observed attributes, current prices and current benchmark results and returns the models that satisfy them — sorted by how many criteria they meet, then best benchmark rank, then release date. No composite score, no single winner.
125 models satisfy at least one of the criteriaNo single winner — these are the observed dimensions; the order is criteria met → best rank → release date.
use case visiondeployment apimax output price 5
| Model | Why | Observed dimensions | Deployments | Ranks | |
|---|---|---|---|---|---|
| DeepSeek-V4.1-FlashOpen weightsDeepSeek |
| 1M context · image · text · reasoning · tool calling · released 10 Sept 2026 · from $0.15 in / $0.60 out · 3 providers |
| livebench-agentic-coding#1livebench-2#5livebench-data-analysis#12+9 | |
| Gemini 3.8 FlashProprietaryGoogle |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 2 Sept 2026 · from $0.375 in / $1.88 out · 2 providers |
| livebench-if#1gpqa-diamond#2mmmu-pro#2+10 | |
| gpt-5.6-solProprietaryOpenAI |
| 1.05M context · image · text · reasoning · tool calling · released 9 Jul 2026 · from $1 in / $5 out · 2 providers |
| terminal-bench#1livebench-coding#3livebench-mathematics#3+13 | |
| Grok 4.3ProprietaryxAI |
| 1M context · document · image · text · reasoning · tool calling · released 30 Apr 2026 · from $1 in / $2 out · 2 providers |
| ifbench#1tau2-bench#9mmmu-pro#30+13 | |
| gpt-5ProprietaryOpenAI |
| 400K context · document · image · text · reasoning · tool calling · released 7 Aug 2025 · from $0.625 in / $5 out · 2 providers |
| aider-polyglot#1swe-bench-verified#17ifbench#42+7 | |
| o3ProprietaryOpenAI |
| 200K context · document · image · text · reasoning · tool calling · released 16 Apr 2025 · from $1 in / $4 out · 2 providers |
| swe-bench-multimodal#1aider-polyglot#4swe-bench-verified#22+8 | |
| Gemma 3 27BOpen weightsGoogle |
| 131.1K context · image · text · tool calling · released 12 Mar 2025 · from $0.08 in / $0.45 out · 2 providers |
| aider-polyglot-well-formed#1aider-polyglot#60scicode#123+7 | |
| GPT-4o-mini (2024-07-18)OpenAI |
| 128K context · document · image · text · tool calling · released 18 Jul 2024 · from $0.15 in / $0.60 out · 2 providers |
| aider-polyglot-well-formed#1aider-polyglot#61 | |
| Gemini 3.7 FlashProprietaryGoogle |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 13 Aug 2026 · from $0.375 in / $1.88 out · 2 providers |
| livebench-if#2mmmu-pro#3scicode#3+10 | |
| o4-miniProprietaryOpenAI |
| 200K context · document · image · text · reasoning · tool calling · released 16 Apr 2025 · from $0.55 in / $2.2 out · 2 providers |
| swe-bench-multimodal#2aider-polyglot#10swe-bench-verified#31+8 | |
| MiniMax-M3Open weightsMiniMax |
| 427B params · 1.05M context · Other · image · text · video · reasoning · tool calling · released 2 Jun 2026 · from $0.30 in / $1.2 out · 4 providers |
| ifbench#3gpqa-diamond#14mmmu-pro#26+13 | |
| gpt-4.1ProprietaryOpenAI |
| 1.05M context · document · image · text · tool calling · released 14 Apr 2025 · from $1 in / $4 out · 2 providers |
| swe-bench-multimodal#3aider-polyglot-well-formed#20aider-polyglot#28+8 | |
| Muse Spark 1.3ProprietaryMeta AI |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 2 Sept 2026 · from $1.25 in / $4.25 out · 1 provider |
| livebench-2#4livebench-if#4scicode#4+10 | |
| DeepSeek V4 Flash Vision ExpOpen weightsDeepSeek |
| 304.6B params · 1.05M context · MIT · image · text · reasoning · tool calling · released 31 Aug 2026 · from $0.11 in / $0.33 out · 3 providers |
| livebench-agentic-coding#4livebench-data-analysis#8livebench-2#21+5 | |
| Step 3.7 FlashOpen weightsStepFun |
| 262.1K context · image · text · video · reasoning · tool calling · released 28 May 2026 · from $0.20 in / $1.15 out · 1 provider |
| tau2-bench#4mmmu-pro#44terminal-bench#54+5 | |
| GLM 5V TurboProprietaryZ.ai (Zhipu AI) |
| 202.8K context · image · text · video · reasoning · tool calling · released 1 Apr 2026 · from $1.2 in / $4 out · 2 providers |
| tau2-bench#4mmmu-pro#65terminal-bench#74+4 | |
| gpt-4oProprietaryOpenAI |
| 128K context · document · image · text · tool calling · released 13 May 2024 · from $1.25 in / $5 out · 2 providers |
| swe-bench-multimodal#4swe-bench-verified#38terminal-bench#176+5 | |
| Qwen3.8 FlashOpen weightsQwen |
| 1M context · image · text · video · reasoning · tool calling · released 26 Aug 2026 · from $0.15 in / $0.47 out · 2 providers |
| livebench-if#5livebench-agentic-coding#10artificial-analysis-intelligence-index#17+10 | |
| gpt-5.6-lunaProprietaryOpenAI |
| 1.05M context · image · text · reasoning · tool calling · released 9 Jul 2026 · from $0.10 in / $0.60 out · 2 providers |
| livebench-coding#5livebench-data-analysis#22scicode#22+10 | |
| Gemini 3.5 FlashProprietaryGoogle |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 19 May 2026 · from $0.75 in / $4.5 out · 2 providers |
| mmmu-pro#5livebench-if#7livebench-language#11+13 | |
| Grok 4.20ProprietaryxAI |
| 1M context · document · image · text · reasoning · tool calling · released 31 Mar 2026 · from $1.25 in / $2.5 out · 2 providers |
| ifbench#5gpqa-diamond#31tau2-bench#41+4 | |
| Gemini 3.6 FlashProprietaryGoogle |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 21 Jul 2026 · from $0.375 in / $1.88 out · 2 providers |
| mmmu-pro#7livebench-if#9livebench-language#13+10 | |
| Muse Spark 1.1ProprietaryMeta AI |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 16 Jul 2026 · from $1.25 in / $4.25 out · 1 provider |
| scicode#7humanitys-last-exam#12livebench-agentic-coding#16+9 | |
| Kimi K2.5Open weightsMoonshot AI |
| 1.03T params · 262.1K context · Other · image · text · reasoning · tool calling · released 1 Jan 2026 · from $0.45 in / $2.25 out · 2 providers |
| swe-bench-multilingual#7swe-bench-verified#10tau2-bench#14+6 | |
| Muse Spark 1.2ProprietaryMeta AI |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 5 Aug 2026 · from $1.25 in / $4.25 out · 1 provider |
| livebench-reasoning#9livebench-if#10scicode#10+9 | |
| Qwen3.6 PlusProprietaryQwen |
| 1M context · image · text · video · reasoning · tool calling · released 2 Apr 2026 · from $0.325 in / $1.95 out · 2 providers |
| tau2-bench#9livebench-coding#24terminal-bench#24+12 | |
| Qwen3.8 27BOpen weightsQwen |
| 27.8B params · 1M context · Apache-2.0 · image · text · video · reasoning · tool calling · released 5 Aug 2026 · from $0.214 in / $2.55 out · 2 providers |
| livebench-agentic-coding#11livebench-if#16livebench-data-analysis#24+10 | |
| gpt-5.4-miniProprietaryOpenAI |
| 400K context · document · image · text · reasoning · tool calling · released 17 Mar 2026 · from $0.375 in / $2.25 out · 2 providers |
| terminal-bench#11scicode#26ifbench#40+13 | |
| Qwen3.5 397B A17BOpen weightsQwen |
| 262.1K context · image · text · video · reasoning · tool calling · released 16 Feb 2026 · from $0.55 in / $3.5 out · 2 providers |
| ifbench#11tau2-bench#18mmmu-pro#32+5 | |
| GLM 5.3 FlashOpen weightsZ.ai (Zhipu AI) |
| 321.3B params · 1.31M context · MIT · image · text · video · reasoning · tool calling · released 25 Aug 2026 · from $0.075 in / $0.25 out · 4 providers |
| artificial-analysis-intelligence-index#12livebench-coding#19livebench-agentic-coding#21+9 | |
| Qwen 3.7 PlusProprietaryQwen |
| 1M context · image · text · reasoning · tool calling · released 3 Jun 2026 · from $0.32 in / $1.28 out · 3 providers |
| ifbench#12mmmu-pro#15terminal-bench#16+5 | |
| Gemini 3 Flash PreviewProprietaryGoogle |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 17 Dec 2025 · from $0.25 in / $1.5 out · 2 providers |
| ifbench#12mmmu-pro#18terminal-bench#40+4 | |
| Claude Sonnet 5ProprietaryAnthropic |
| 1M context · image · text · reasoning · tool calling · released 30 Jun 2026 · from $1 in / $5 out · 2 providers |
| livebench-coding#13livebench-agentic-coding#15livebench-mathematics#15+10 | |
| gpt-5-miniProprietaryOpenAI |
| 400K context · document · image · text · reasoning · tool calling · released 7 Aug 2025 · from $0.125 in / $1 out · 2 providers |
| swe-bench-multilingual#13swe-bench-verified#24ifbench#30+7 | |
| Kimi K2.6Open weightsMoonshot AI |
| 1.03T params · 262.1K context · Other · image · text · reasoning · tool calling · released 14 Apr 2026 · from $0.95 in / $4 out · 3 providers |
| tau2-bench#14mmmu-pro#21ifbench#22+13 | |
| gpt-5.1ProprietaryOpenAI |
| 400K context · document · image · text · reasoning · tool calling · released 13 Nov 2025 · from $0.625 in / $5 out · 2 providers |
| swe-bench-verified#15terminal-bench#21mmmu-pro#39+5 | |
| Gemini 3.1 Flash-Lite PreviewProprietaryGoogle |
| 1.05M context · audio · document · image · text · video · reasoning · tool calling · released 3 Mar 2026 · from $0.25 in / $1.5 out · 2 providers |
| ifbench#16mmmu-pro#38scicode#66+5 | |
| Llama 4 MaverickOpen weightsMeta AI |
| 1.05M context · image · text · tool calling · released 5 Apr 2025 · from $0.20 in / $0.696 out · 1 provider |
| aider-polyglot-well-formed#17aider-polyglot#54mmmu-pro#101+7 | |
| Qwen3.6 35B A3BOpen weightsQwen |
| 36B params · 262.1K context · Apache-2.0 · image · text · video · reasoning · tool calling · released 15 Apr 2026 · from $0.10 in / $0.90 out · 1 provider |
| tau2-bench#22mmmu-pro#45terminal-bench#59+5 | |
| gpt-5.4-nanoProprietaryOpenAI |
| 400K context · document · image · text · reasoning · tool calling · released 17 Mar 2026 · from $0.10 in / $0.625 out · 2 providers |
| livebench-mathematics#22ifbench#23terminal-bench#30+13 |
Green chips = criteria satisfied by an observed fact; dashed amber = satisfied by an estimate (hardware fit). Ranks are positions inside each benchmark's primary comparability group.
MethodDeterministic filters over observed attributes, current prices and current benchmark results of canonical models; sorted by number of satisfied criteria, then best benchmark rank, then release date. No composite score. Hardware fit is an estimate. /methodology