SWE-bench Verified (mini-swe-agent-v2)

A display-only SWE-bench Verified reference from Arcee AI's Trinity-Large-Thinking comparison chart.

  • In index
  • Coding
ModelsModels with a result.5
Top resultBest result on this test.75.6%Claude Opus 4.6
Top-3 spreadPoints from first to third.2.8 pts
YearYear of release.N/A

Result and price

020406080100$0.3$1$3$10$30RESULTOUTPUT PRICE PER 1M TOKENS · LOG SCALE

Results

5 results
#ModelResultThe score from the source.
03GLM-572.8%