ByteDance
partialShows if the model has enough results for an index.Seed 2.1 Turbo
Seed 2.1 Turbo is a reasoning model from ByteDance in the Seed 2.1 family. 6 benchmarks count toward its score, in 4 categories.
IndexOverall score out of 100.Unranked
CoverageShare of the index weight with results.65%
SpeedOutput tokens per second.42/s
Input / 1MUS dollars per 1M input tokens.$0.5
Output / 1MUS dollars per 1M output tokens.$2.5
ContextMaximum tokens in one request.262K
EloLMArena rating and rank.N/A
The index is a score out of 100. The ± range shows how much it can change.
CapabilitiesScore per category, out of 100.
Out of 100Results
6 counted| BenchmarkThe test name. | CategoryThe capability that the test measures. | ResultThe score from the publisher. | IndexThis result as a score out of 100. | RunThe settings of the run. | DateDate of the result. | Published byThe source of the result. |
|---|---|---|---|---|---|---|
| MathVision | Multimodal | 90.1% | — | — | — | Qwen |
| Video-MME | Multimodal | 89.0% | — | — | — | Video-MME benchmark team |
| CharXiv Reasoning | Multimodal | 82.5% | 60.0 | — | — | CharXiv authors |
| Massive Multi-discipline Multimodal Understanding Pro | Multimodal | 80.1% | 57.7 | — | — | MMMU-Pro authors |
| ERQA | Multimodal | 71.3% | 65.6 | — | — | Qwen |
| Terminal-Bench 2.1 (provider run) | Agentic | 67.6% | 63.9 | — | — | DeepSeek-AI |
| Terminal-Bench 2.1 (provider run) | Agentic | 67.6% | 63.9 | — | — | DeepSeek-AI |
| SuperGPQA: Scaling LLM Evaluation Across 285 Graduate Disciplines | Knowledge | 67.4% | 53.3 | — | — | Xiaoxuan Du et al. |
| BabyVision | Multimodal | 62.9% | — | — | — | Meta AI |
| NL2Repo | Coding | 43.7% | 59.2 | — | — | MiniMax |
| ZeroBench | Multimodal | 11.0% | — | — | — | Meta AI |
6 benchmarks count, from 7 of 11 results. A grey row does not count. Too few models took that benchmark.