A family of translation tasks configured in lm-evaluation-harness.
unassessed
| Category | translation |
|---|---|
| Subcategory | machine translation |
| Page status | active |
| Metric | BLEU, TER, and chrF (reported together for the WMT14/WMT16 groups; other groups may differ) |
| Direction | higher_is_better |
| Publisher | EleutherAI |
The group covers translation evaluations including WMT14, WMT16, WMT20, and IWSLT2017 task groups.
Generate a translation for a source sentence.
No model card in ModelSpec reports this benchmark yet.