Catalan translation and cultural adaptation of EQ-Bench version 2's 167-171 question emotion-intensity rating task, released by the Barcelona Supercomputing Center.
unassessed
| Category | reasoning |
|---|---|
| Subcategory | emotional and social intelligence (dialogue emotion-intensity prediction), Catalan translation |
| Page status | active |
| Metric | EQ-Bench score (distance from reference ratings) |
| Direction | higher_is_better |
| Unit | points |
| Dataset size | 167 |
| Dataset licence | CC BY 4.0 |
| Publisher | Barcelona Supercomputing Center (BSC), Language Technologies Unit |
eq_bench_ca shows a model a short dialogue translated and culturally adapted into Catalan, then asks it to rate the intensity (0-10) of four named emotions one character is likely feeling at the end of the scene -- the same task format as English EQ-Bench. The Barcelona Supercomputing Center's Language Technologies Unit produced the adaptation: converting adjectival emotion labels to nominal forms to avoid Catalan grammatical-gender ambiguity, replacing Anglo-Saxon character names with Catalan ones (with correct definite articles), and unifying emotion labels that were equivalent but differently inflected in the English original.
Given a Catalan dialogue and four named emotions, output an intensity rating from 0 to 10 for each in a fixed format; scored by the same distance-from-reference formula as EQ-Bench v2 (see the family page). lm-evaluation-harness runs it as task `eqbench_ca`, generating greedily at temperature 0.
No model card in ModelSpec reports this benchmark yet.