Coding model rankings

Pick a model for the work

Compare publisher-reported coding results by benchmark. No composite score. Missing results stay missing.

25 models22 routableRevised 2026-09-29

DeepSWE v1.1

13 models · publisher-reported scores

Higher scores rank first. Publisher test setups vary. Prices blend input and output 3:1; missing results are excluded.

Compare score against price

Scroll to compare the full chart →

Method, sources, and raw data

What is ranked

Publisher-reported scores. Compare within each benchmark.

What a gap means

The publisher did not report that result. MiniRouter does not estimate it or turn it into zero.

How to compare

Harnesses, scaffolds, and effort settings may differ even when vendors use the same benchmark name.