Gender Sensitivity Test - English

An English BIG-bench task that scores occupation-title gender bias, gender identification from names and terms, and PTB perplexity.

Also known as: gender sensitivity test English

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categorysafety
SubcategoryBIG-bench English occupation-bias, gender-identification, and PTB probes
Page statusunknown
Metricnormalized_aggregate_score (BIG-bench composite of bias, stereotype, identification, PTB)
Directionhigher_is_better
Dataset size2035
Dataset licenceApache-2.0
PublisherGoogle (BIG-bench collaboration)

What it measures

gender_sensitivity_english runs three zero-shot tests. Neutrality measures whether next-token probabilities after a gender-neutral occupation plus "is" favour male, female, or listed non-binary terms, and whether occupation distributions given those terms differ (stereotype). Identification measures whether the same comparison recovers labelled gender from gendered nouns and SSA-filtered given names. A third arm reports character-level negative perplexity on a Penn Treebank string shipped in test_data.json, because de-biasing can hurt language modelling.

Task format

Programmatic next-token comparison. Neutrality and identification prompts end in " is ". Preferred scores: gender_bias_score, gender_minority_bias_score, gender_stereotype_score and gender_minority_stereotype_score on [-1, 0]; mean_accuracy on identification; negative_perplexity on PTB. Canary GUID embedded. Zero-shot.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub