DeepSeek R1 Qwen3 8B

DeepSeek R1 Qwen3 8B is a model from DeepSeek. It has 3 scored results in 3 categories. It has no index.

  • Input: text
partialShows if the model has enough results for an index.
IndexOverall score out of 100.UnrankedMedian 50.7
CoverageShare of the index weight with results.40%
SpeedOutput tokens per second.—Median 62/s
First tokenSeconds to the first output token.N/AMedian 1.6 s
Input / 1MUS dollars per 1M input tokens.N/AMedian $0.6
Output / 1MUS dollars per 1M output tokens.N/AMedian $2.5
ContextMaximum tokens in one request.N/AMedian 262K
EloLMArena rating and rank.N/AMedian 1427

Capabilities

Score per category, out of 100.
Out of 100
AgenticMulti-step tasks with tools.None countedN/A—
CodingCode writing and repair.None countedN/A—
ReasoningLogic problems and puzzles.1 counted26.0—
MultimodalTasks with images and text.None countedN/A—
KnowledgeFacts and expert knowledge.1 counted5.0—
MultilingualTasks in many languages.None countedN/A—
InstructionTasks with strict rules in the prompt.None countedN/A—
MathMath problems.1 counted34.8—

Results

3 counted
BenchmarkThe test name.ResultThe score from the source.PlacePlace among results on this benchmark.
OTIS Mock AIME 2024-2025Math43.9% ±6.8112 / 185
Chess PuzzlesReasoning3.0% ±1.799 / 136
GPQA diamond · Epoch AIKnowledge9.3% ±1.1203 / 203

Sources

BenchLM benchmark aggregationCC BY-NC 4.0 · Data from BenchLM.aiEpoch AI, collected directlyCC BY — free to use and redistribute with attribution

More from DeepSeek

All models
DeepSeek V4 Pro 081367.5DeepSeek V4.1 Flash67.0DeepSeek V4 Flash 073163.0DeepSeek V4 Pro 042358.5DeepSeek V3.2 (Thinking)49.1