Apex
A high-difficulty mathematical reasoning benchmark reported in DeepSeek-V4 model evaluations.
ModelsModels with a result.8
Top resultBest result on this test.N/A
Top-3 spreadPoints from first to third.N/A
YearYear of release.N/A
Result and price
Results
8 results#ModelResultThe score from the source.