WebQuestions evaluates short-answer questions whose answers are grounded in Freebase entities.
unassessed
| Category | knowledge |
|---|---|
| Subcategory | question answering |
| Page status | active |
| Metric | exact match |
| Direction | higher_is_better |
| Unit | percent |
| Dataset size | 2032 |
| Publisher | Stanford NLP Group |
WebQuestions tests open-domain question answering from natural-language questions. The lm-evaluation-harness task identifies the WebQuestions task and its answer-extraction protocol.
Natural-language question followed by a short free-form answer.
No model card in ModelSpec reports this benchmark yet.