Skip to content
AI Atlas
BenchmarkActivecategory · agenticfamily · terminal-bench · variant 2.0

Terminal-Bench 2.0

tbench.ai

revised, harder set of terminal tasks (Terminal-Bench 2.0)

data quality57

Updated 49 min ago · first seen 12 Sept 2026

Metric
accuracy · %
Current results
0
Models
0
Current leader

No results recorded for this benchmark yet — its sources are being connected. The definition, aliases and variants are kept so links resolve; nothing is fabricated.

Leaderboard 0 models

Select models with +, then Compare.

No result in this group with these filters

Relax the trust / organization filters or pick another comparability group.