Social Support

A BIG-bench task that asks a model to classify a comment from an online support community as supportive, neutral or unsupportive of the post it replies to.

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categorysafety
Subcategorysocial and emotional understanding: classify supportiveness of a comment in an online conversation
Page statusactive
MetricMacro-F1 against majority-vote crowd labels (BIG-bench also records multiple_choice_grade)
Directionhigher_is_better
UnitF1
Dataset size897
PublisherGoogle (BIG-bench collaboration); task authors Zijian Wang and David Jurgens

What it measures

Social Support gives a model a post from an online community and a reply comment, and asks the model to judge whether the reply is supportive, neutral or unsupportive of the poster. The annotated corpus behind the task spans Reddit, StackExchange and Wikipedia talk-page interactions, not a single platform. It targets a narrow slice of social and emotional understanding: recognising encouragement, empathy or advice versus dismissiveness or hostility in short, informal, emotionally loaded text, not general sentiment polarity.

Task format

Zero-shot multiple-choice classification. Each of 897 examples presents a post/reply pair and asks the model to pick one of three labels (supportive, neutral, unsupportive); BIG-bench scores it via the multiple_choice_grade metric over the three answer options.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub