FewCLUE: BUSTM (Dialogue Short Text Matching)

FewCLUE's dialogue short-text matching task: judge whether two short colloquial Chinese sentences share the same intent, learned from 32 labelled training pairs.

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categoryreasoning
Subcategorydialogue short-text semantic matching (binary intent match), few-shot
Page statusunknown
Metricaccuracy
Directionhigher_is_better
Unit%
Dataset size1772
PublisherCLUE team

What it measures

Two short, colloquial Chinese sentences drawn from a voice assistant's intent-matching logs; the model judges whether they express the same intent, a binary match/no-match decision, learned few-shot from 32 labelled training pairs.

Task format

Sentence-pair binary classification (match / no match), graded on the single correct label; evaluated from a 32-example few-shot training split, one of five parallel splits FewCLUE provides for this task.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub