Skip to content
AI Atlas

SWE-bench Lite — cost vs performance

Best current row per canonical model in the group “resolved · board=Lite · system=reproducedRG” (1 models) against ESTIMATED memory at 4-bit, 8K context (GB). The dashed line is the Pareto frontier: no model is both better and cheaper than a point on it.

XOutput priceInput priceParametersContextMemory (est.)ScaleloglinearBubblecontextparamsnone
Group

No model has both a score in this group and a value on this axis

Try another axis or comparability group.