Alibaba
partialShows if the model has enough results for an index.Qwen3.5 397B A17B
Qwen3.5 397B A17B is a model from Alibaba. 6 benchmarks count toward its score, in 4 categories.
IndexOverall score out of 100.Unranked
CoverageShare of the index weight with results.60%
SpeedOutput tokens per second.12/s
Input / 1MUS dollars per 1M input tokens.$0.55
Output / 1MUS dollars per 1M output tokens.$3.5
ContextMaximum tokens in one request.262K
EloLMArena rating and rank.1438 (#71)
The index is a score out of 100. The ± range shows how much it can change.
77,007 votes. Elo shows what people prefer. It does not change the score.
CapabilitiesScore per category, out of 100.
Out of 100Results
6 counted| BenchmarkThe test name. | CategoryThe capability that the test measures. | ResultThe score from the publisher. | IndexThis result as a score out of 100. | RunThe settings of the run. | DateDate of the result. | Published byThe source of the result. |
|---|---|---|---|---|---|---|
| τ²-bench Telecom | Agentic | 97.8% | 69.0 | enabled effort · Sierra | 2 Mar 2026 | Sierra Research |
| GPQA diamond | Knowledge | 86.4% | 58.0 | none effort | — | Epoch AI |
| τ²-bench Retail | Agentic | 84.4% | 59.4 | enabled effort · Sierra | 30 Apr 2026 | Sierra Research |
| OTIS Mock AIME 2024-2025 | Math | 82.2% | 57.5 | none effort | — | Epoch AI |
| τ²-bench Airline | Agentic | 81.5% | 57.3 | enabled effort · Sierra | 2 Mar 2026 | Sierra Research |
| FrontierMath-Tiers-1-3-v2-Private | Math | 31.2% | 49.0 | none effort | — | Epoch AI |
| Mystery Game Puzzles | Reasoning | 18.0% | 53.1 | none effort | — | Epoch AI |
| Chess Puzzles | Reasoning | 13.0% | 40.3 | none effort | — | Epoch AI |
| τ²-bench Banking | Agentic | 9.8% | 5.9 | enabled effort · Sierra | 4 Aug 2026 | Sierra Research |
6 benchmarks count, from 9 of 9 results. A grey row does not count. Too few models took that benchmark.