TriviaQA reading comprehension evaluated through the OpenCompass TriviaQArc configuration.
unassessed
| Category | knowledge |
|---|---|
| Subcategory | open-domain question answering |
| Page status | active |
| Metric | exact match |
| Direction | higher_is_better |
| Unit | percent |
| Publisher | OpenCompass |
The task presents a question with evidence and asks the model to produce the answer.
Evidence passage, question, and generated answer.
No model card in ModelSpec reports this benchmark yet.