Apex

A high-difficulty mathematical reasoning benchmark reported in DeepSeek-V4 model evaluations.

  • Not in index
  • Math
ModelsModels with a result.8
Top resultBest result on this test.N/A
Top-3 spreadPoints from first to third.N/A
YearYear of release.N/A

Result and price

020406080100$0.1$0.3$1$3$10RESULTOUTPUT PRICE PER 1M TOKENS · LOG SCALE

Results

8 results
#ModelResultThe score from the source.
A.X K245.8%
Qwen3.7 Max44.5%
ZAYA1-8B32.2%
Qwen3.7 Plus22.7%