Z.AI
availableShows if the model has enough results for an index.GLM-4.5
GLM-4.5 is a non-reasoning model from Z.AI. 8 benchmarks count toward its score, in 4 categories.
IndexOverall score out of 100.46.2 ±9.2
CoverageShare of the index weight with results.65%
SpeedOutput tokens per second.15/s
Input / 1MUS dollars per 1M input tokens.$0.6
Output / 1MUS dollars per 1M output tokens.$2.2
ContextMaximum tokens in one request.131K
EloLMArena rating and rank.1430 (#85)
The index is a score out of 100. The ± range shows how much it can change.
23,712 votes. Elo shows what people prefer. It does not change the score.
CapabilitiesScore per category, out of 100.
Out of 100Results
8 counted| BenchmarkThe test name. | CategoryThe capability that the test measures. | ResultThe score from the publisher. | IndexThis result as a score out of 100. | RunThe settings of the run. | DateDate of the result. | Published byThe source of the result. |
|---|---|---|---|---|---|---|
| MATH 500 | Math | 94.0% | 46.9 | — | 9 Jan 2026 | Vals AI |
| MGSM | Multilingual | 90.8% | — | — | 9 Jan 2026 | Vals AI |
| AIME | Math | 86.7% | 54.1 | — | 16 Apr 2026 | Vals AI |
| MMLU Pro | Knowledge | 81.2% | 48.5 | — | 1 Sept 2026 | Vals AI |
| GPQA Diamond | Knowledge | 72.2% | 44.9 | — | 1 Sept 2026 | Vals AI |
| LiveCodeBench | Coding | 67.4% | 45.6 | — | 1 Sept 2026 | Vals AI |
| SWE-bench Verified | Coding | 59.2% | 45.0 | Undisclosed | 1 Sept 2026 | SWE-bench team |
| SWE-bench Verified | Coding | 54.2% | 41.0 | mini-SWE-agent | 26 Feb 2026 | SWE-bench team |
| Terminal-Bench 1.0 | Agentic | 41.3% | 46.5 | — | 12 Jan 2026 | Vals AI |
| IOI v1 | Coding | 2.9% | 41.6 | — | 9 Aug 2026 | Vals AI |
8 benchmarks count, from 9 of 10 results. A grey row does not count. Too few models took that benchmark.