FewCLUE's dialogue short-text matching task: judge whether two short colloquial Chinese sentences share the same intent, learned from 32 labelled training pairs.
unassessed
| Category | reasoning |
|---|---|
| Subcategory | dialogue short-text semantic matching (binary intent match), few-shot |
| Page status | unknown |
| Metric | accuracy |
| Direction | higher_is_better |
| Unit | % |
| Dataset size | 1772 |
| Publisher | CLUE team |
Two short, colloquial Chinese sentences drawn from a voice assistant's intent-matching logs; the model judges whether they express the same intent, a binary match/no-match decision, learned few-shot from 32 labelled training pairs.
Sentence-pair binary classification (match / no match), graded on the single correct label; evaluated from a 32-example few-shot training split, one of five parallel splits FewCLUE provides for this task.
No model card in ModelSpec reports this benchmark yet.