Mistral Large 2411

Mistral Large 2411 is a model from Mistral. It has 5 scored results in 3 categories. It has no index.

  • Input: text
partialShows if the model has enough results for an index.
IndexOverall score out of 100.UnrankedMedian 50.7
CoverageShare of the index weight with results.45%
SpeedOutput tokens per second.—Median 62/s
First tokenSeconds to the first output token.N/AMedian 1.6 s
Input / 1MUS dollars per 1M input tokens.N/AMedian $0.6
Output / 1MUS dollars per 1M output tokens.N/AMedian $2.5
ContextMaximum tokens in one request.N/AMedian 262K
EloLMArena rating and rank.1265±4 #270Median 1427

Capabilities

Score per category, out of 100.
Out of 100
AgenticMulti-step tasks with tools.None countedN/A—
CodingCode writing and repair.1 counted18.0—
ReasoningLogic problems and puzzles.None countedN/A—
MultimodalTasks with images and text.None countedN/A—
KnowledgeFacts and expert knowledge.2 counted24.7—
MultilingualTasks in many languages.None countedN/A—
InstructionTasks with strict rules in the prompt.None countedN/A—
MathMath problems.2 counted20.8—

Results

5 counted
BenchmarkThe test name.ResultThe score from the source.PlacePlace among results on this benchmark.
MATH 500Math74.4%45 / 59
MMLU ProKnowledge69.7%119 / 135
GPQA Diamond · Vals AIKnowledge47.7%123 / 135
LiveCodeBenchCoding37.1%132 / 142
AIMEMath9.2%90 / 95
MGSMMultilingual87.2%60 / 74
MedQAHealthcare76.2%78 / 95
TaxEval v2Finance63.8%118 / 143

Sources

BenchLM benchmark aggregationCC BY-NC 4.0 · Data from BenchLM.aiVals AI, collected directlyNo licence stated. Read from the public leaderboard and credited to Vals AI

More from Mistral

All models
Mistral Large 462.1Mistral Medium 3.5 128B49.1Mistral Medium 3.541.9Magistral Medium38.0Mistral Small 436.5