Private Stanford Health Care MedHELM task: from a clinical note, answer whether an ENT referral is supported, with A yes, B no, or C no mention.
unassessed
| Category | domain |
|---|---|
| Subcategory | ear, nose, and throat specialist referral from notes |
| Page status | active |
| Metric | exact_match |
| Direction | higher_is_better |
| Unit | 0-1 scale |
| Dataset size | 1000 |
| Publisher | Stanford Health Care and Stanford CRFM (MedHELM) |
ENT-Referral tests whether a model can read an unstructured English clinical note and decide if the note supports referring the patient to an ear, nose, and throat (ENT) specialist. Unlike the binary Stanford Health Care tasks, HELM allows a third label, C, for no mention of referral. The census id is shc_ent; the runnable HELM scenario is shc_ent_med.
Joint multiple-choice generation. HELM prefixes a running item counter, then asks for A (yes), B (no), or C (no mention), with no extra text.
No model card in ModelSpec reports this benchmark yet.