Humanity's Last Exam without tools

Tool-free variant of Humanity's Last Exam that isolates a model's raw frontier reasoning.

  • In index
  • Knowledge
ModelsModels with a result.38
Top resultBest result on this test.64.4%Claude Opus 5.5
Top-3 spreadPoints from first to third.5.4 pts
YearYear of release.N/A

Result and price

020406080100$0.3$1$3$10$30$100$300RESULTOUTPUT PRICE PER 1M TOKENS · LOG SCALE

Results

38 results
#ModelResultThe score from the source.
13Kimi K343.5%
17Muse Spark42.8%
19GPT-5.541.4%
20GLM-5.240.5%
22GPT-5.439.8%
25Grok 4.2031.6%
28Inkling30.0%