Benchmark

MMLU French

French language variant of the Massive Multitask Language Understanding benchmark, evaluating language models across 57 tasks including elementary…

Modality
text
Categories
legal, math, reasoning, language, finance, general, healthcare
Openness
unknown
Source
llm_stats
Reported scores
1

Reported scores

llm_stats

ModelOrganizationReported valueReportedEvidence
Mistral Large 2Mistral0.8282024-07-24evidence

Scores are partitioned by the source that reported them and are never merged into a single cross-source ranking, because the sources measure different things and say so.