MMLU clinical African languages (HELM)

HELM wrap of human-translated MMLU clinical knowledge, college medicine, and virology items in 11 African languages, scored by exact match.

Also known as: mmlu_cm_ck_vir, Bridging-the-Gap MMLU-Clinical

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categoryknowledge
Subcategoryhuman-translated MMLU clinical subjects in 11 African languages
Page statusunknown
Metricexact_match
Directionhigher_is_better
Dataset size265
Dataset licenceMIT (translation release; MMLU source also MIT)
PublisherInstitute for Disease Modeling, Bill & Melinda Gates Foundation, and Ghamut Corporation; HELM wrap by Stanford CRFM

What it measures

mmlu_clinical_afr is HELM's multiple-choice wrap of three MMLU health subjects translated into 11 African languages. Each item is a four-option question in the target language. Default constructor arguments are subject clinical_knowledge and lang af (Afrikaans). The run spec also accepts college_medicine and virology, and ISO codes af, zu, xh, am, bm, ig, nso, sn, st, tn, ts. It is text-only. It is not English MMLU, not MMLU-ProX, and not Global-MMLU.

Task format

Joint multiple-choice. Instruction: "The following are multiple choice questions (with answers) about {subject} in {language}." Input noun Question, output noun Answer. Adapter default max_train_instances is 5; HELM maps the 5-item dev csv to TRAIN_SPLIT, so the usual protocol is 5-shot from dev. Main metric exact_match on test.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub