COVIDDialog (HELM English medical dialogue)

HELM generation task: read an English COVID-19 patient question and write the doctor's reply, scored with overlap metrics.

Also known as: COVIDDialog, covid_dialogue, CovidDialog

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categorydomain
SubcategoryHELM English patient-to-doctor response generation on COVID-19 consultations
Page statusunknown
Metricopen-ended overlap (exact_match, quasi_exact_match, f1_score, rouge_l, bleu_1, bleu_4)
Directionhigher_is_better
Dataset size603
PublisherUCSD AI4H (dataset); Stanford CRFM (HELM scenario)

What it measures

covid_dialog is HELM's wrap of the English CovidDialog consultations, not a new item set. The model reads a patient's COVID-19 or pneumonia concern and must write the doctor's reply. HELM strips a leading "patient: " from the source line and prompts with Patient / Doctor. The original English collection is described as 603 consultations with id, URL, condition description, and dialogue. This is English doctor-response generation, not the Chinese Haodf.com dump and not HELM med_dialog.

Task format

Instruction "Generate a response given a patient's questions and concerns." then Patient: … / Doctor: … . Default five in-context examples, temperature 0, max_tokens 128. Run spec name covid_dialog; group COVIDDialog.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub