BIG-bench Topical-Chat evaluates open-domain response generation in conversations grounded in topical information.
unassessed
| Category | generation |
|---|---|
| Subcategory | open-domain dialogue |
| Page status | active |
| Metric | BLEU |
| Direction | higher_is_better |
| Unit | score |
| Dataset size | 22295 |
| Publisher | Google BIG-bench |
The task asks a model to generate a response in an open-domain conversation. Its examples contain multi-turn dialogue about topics such as people, science, sports, and culture.
Dialogue context followed by a generated response.
No model card in ModelSpec reports this benchmark yet.