Benchmark
WMT23
The Eighth Conference on Machine Translation (WMT23) benchmark evaluating machine translation systems across 8 language pairs (14 translation directions)…
- Modality
- text
- Categories
- language, healthcare
- Openness
- unknown
- Source
- llm_stats
- Reported scores
- 4
Reported scores
llm_stats
| Model | Organization | Reported value | Reported | Evidence |
|---|---|---|---|---|
| Gemini 1.0 Pro | 0.717 | 2024-02-15 | evidence | |
| Gemini 1.5 Flash | 0.741 | 2024-05-01 | evidence | |
| Gemini 1.5 Flash 8B | 0.726 | 2024-03-15 | evidence | |
| Gemini 1.5 Pro | 0.751 | 2024-05-01 | evidence |
Scores are partitioned by the source that reported them and are never merged into a single cross-source ranking, because the sources measure different things and say so.