WebQuestions

WebQuestions evaluates short-answer questions whose answers are grounded in Freebase entities.

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categoryknowledge
Subcategoryquestion answering
Page statusactive
Metricexact match
Directionhigher_is_better
Unitpercent
Dataset size2032
PublisherStanford NLP Group

What it measures

WebQuestions tests open-domain question answering from natural-language questions. The lm-evaluation-harness task identifies the WebQuestions task and its answer-extraction protocol.

Task format

Natural-language question followed by a short free-form answer.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub