EQ-Bench (Catalan)

Catalan translation and cultural adaptation of EQ-Bench version 2's 167-171 question emotion-intensity rating task, released by the Barcelona Supercomputing Center.

Also known as: EQ-bench_ca

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categoryreasoning
Subcategoryemotional and social intelligence (dialogue emotion-intensity prediction), Catalan translation
Page statusactive
MetricEQ-Bench score (distance from reference ratings)
Directionhigher_is_better
Unitpoints
Dataset size167
Dataset licenceCC BY 4.0
PublisherBarcelona Supercomputing Center (BSC), Language Technologies Unit

What it measures

eq_bench_ca shows a model a short dialogue translated and culturally adapted into Catalan, then asks it to rate the intensity (0-10) of four named emotions one character is likely feeling at the end of the scene -- the same task format as English EQ-Bench. The Barcelona Supercomputing Center's Language Technologies Unit produced the adaptation: converting adjectival emotion labels to nominal forms to avoid Catalan grammatical-gender ambiguity, replacing Anglo-Saxon character names with Catalan ones (with correct definite articles), and unifying emotion labels that were equivalent but differently inflected in the English original.

Task format

Given a Catalan dialogue and four named emotions, output an intensity rating from 0 to 10 for each in a fixed format; scored by the same distance-from-reference formula as EQ-Bench v2 (see the family page). lm-evaluation-harness runs it as task `eqbench_ca`, generating greedily at temperature 0.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub