PrivacyDetection

Private MedHELM task: decide whether a patient-portal message contains confidential or privacy-leaking information, answering A for yes or B for no.

Also known as: shc_privacy_med, PrivacyDetection

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categorydomain
Subcategoryconfidential content in patient portal messages
Page statusactive
Metricexact_match
Directionhigher_is_better
Unit0-1 scale
Dataset size300
PublisherStanford University Department of Pediatrics, Stanford Health Care, and Stanford CRFM (MedHELM)

What it measures

PrivacyDetection tests whether a model can read an English patient-portal message and say whether it contains confidential or privacy-leaking information that should be protected. HELM metadata describes messages from patients or caregivers. The census id is shc_privacy; the runnable HELM scenario is shc_privacy_med.

Task format

Joint multiple-choice generation. HELM instructs the model to review clinical messages for confidential information and answer A for yes or B for no.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub