Inference
Explore
Leaderboard
Benchmarks
Methodology
24 Sept 2026
Benchmarks
/
cursorBench31
cursorBench31
Not in index
Coding
Models
?
Models with a result.
7
Top result
?
Best result on this test.
N/A
Top-3 spread
?
Points from first to third.
N/A
Year
?
Year of release.
N/A
Result and price
0
20
40
60
80
100
$1
$3
$10
$30
$100
RESULT
PARETO
OUTPUT PRICE PER 1M TOKENS · LOG SCALE
Results
7 results
#
Model
Result
?
The score from the source.
—
Claude Fable 5
70.6
%
—
CU
Composer 2.5
63.2
%
—
GPT-5.5
59.2
%
—
Claude Opus 4.8
58.4
%
—
Gemini 3.5 Flash
49.8
%
—
Claude Sonnet 4.6
48.8
%
—
MA
Kimi K2.6
47.6
%
cursorBench31 results — Inference360