EuroLLM 9B Instruct 2512

EuroLLM 9B Instruct 2512 is a model from Unknown. It has 1 scored result in 1 category. It has no index.

  • Input: text
partialShows if the model has enough results for an index.
IndexOverall score out of 100.UnrankedMedian 50.7
CoverageShare of the index weight with results.5%
SpeedOutput tokens per second.—Median 62/s
First tokenSeconds to the first output token.N/AMedian 1.6 s
Input / 1MUS dollars per 1M input tokens.N/AMedian $0.6
Output / 1MUS dollars per 1M output tokens.N/AMedian $2.5
ContextMaximum tokens in one request.N/AMedian 262K
EloLMArena rating and rank.N/AMedian 1427

Capabilities

Score per category, out of 100.
Out of 100
AgenticMulti-step tasks with tools.None countedN/A—
CodingCode writing and repair.None countedN/A—
ReasoningLogic problems and puzzles.None countedN/A—
MultimodalTasks with images and text.None countedN/A—
KnowledgeFacts and expert knowledge.None countedN/A—
MultilingualTasks in many languages.1 counted58.1—
InstructionTasks with strict rules in the prompt.None countedN/A—
MathMath problems.None countedN/A—

Results

1 counted
BenchmarkThe test name.ResultThe score from the source.PlacePlace among results on this benchmark.
EuroEval PortugueseMultilingual46.1% ±0.680 / 180
EuroEval DutchMultilingual47.1% ±0.675 / 167
EuroEval ItalianMultilingual45.0% ±0.779 / 172
EuroEval FrenchMultilingual49.7% ±0.584 / 176
EuroEval SwedishMultilingual42.2% ±0.995 / 186
EuroEval SpanishMultilingual35.4% ±0.789 / 174
EuroEval GermanMultilingual35.5% ±0.696 / 182
EuroEval PolishMultilingual39.1% ±0.7116 / 205

Sources

BenchLM benchmark aggregationCC BY-NC 4.0 · Data from BenchLM.aiEuroEval, collected directlyMIT — the leaderboard site and its CSV routes are in the licensed repository