Inference
Explore
Leaderboard
Benchmarks
Methodology
24 Sept 2026
Benchmarks
/
cursorBench40
cursorBench40
Not in index
Coding
Models
?
Models with a result.
6
Top result
?
Best result on this test.
N/A
Top-3 spread
?
Points from first to third.
N/A
Year
?
Year of release.
N/A
Result and price
0
20
40
60
80
100
$3
$10
$30
$100
RESULT
PARETO
OUTPUT PRICE PER 1M TOKENS · LOG SCALE
Results
6 results
#
Model
Result
?
The score from the source.
—
Claude Opus 5.5
57.8
%
—
Claude Fable 5.1
51.8
%
—
Grok 4.7
46.3
%
—
GPT-5.6 Sol
41.7
%
—
Gemini 3.8 Flash
39.6
%
—
Claude Sonnet 5
34.1
%
cursorBench40 results — Inference360