EuroEval
EuroEval Swedish
How well a model handles Swedish, as the unweighted mean of EuroEval's primary metric over its 8 Swedish datasets (swerec, suc3, scala_sv, multi_wiki_qa_sv, swedn, skolprov, swedish_facts, winogrande_sv). The primary metric differs by task — MCC for classification, knowledge and common-sense reasoning, micro-F1 without MISC for entity recognition, F1 for reading comprehension, chrF++ for summarization, METEOR for simplification — and is always the first of the two the board prints. Only models measured on every dataset are included, so each mean covers the same ground. Most rows are the validation split; some older open-weight rows are the test split, which EuroEval ranks in the same table.
ModelsModels with a result.173
Top-3 spreadPoints from first to third.1.7 pts
YearYear of release.N/A
Result and price
Results
181 results#ModelResultThe score from the source.