Egyptian Arabic four-way sentence completion translated from HellaSwag; lm-eval scores the 10,042-item validation split.
unassessed
| Category | reasoning |
|---|---|
| Subcategory | Egyptian Arabic commonsense sentence completion (translated HellaSwag) |
| Page status | active |
| Metric | accuracy (acc); also acc_norm |
| Direction | higher_is_better |
| Unit | % |
| Dataset size | 10052 |
| Dataset licence | MIT |
| Publisher | UBC-NLP (University of British Columbia) |
EgyHellaSwag is a machine-translated Egyptian Arabic (Masri / ISO arz) version of HellaSwag. Each item gives an activity label and a context sentence plus four endings. The model must pick the plausible continuation. It tests dialectal commonsense sentence completion, not Egyptian cultural knowledge written from scratch. Items keep original HellaSwag source_id values (ActivityNet and WikiHow).
Four-way multiple choice. lm-eval task egyhellaswag concatenates activity_label and ctx as the query, uses endings as choices, and scores the integer label. Metrics: acc and acc_norm. Training split is 10 rows; scoring uses validation.
No model card in ModelSpec reports this benchmark yet.