Moral Permissibility

A 342-item BIG-bench task that asks yes or no whether an action in a short moral-dilemma story is permissible, using majority human labels.

Also known as: Judging Moral Permissibility

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categoryreasoning
SubcategoryBIG-bench trolley-style yes/no moral dilemmas (342 items)
Page statusunknown
Metricmultiple_choice_grade
Directionhigher_is_better
Unit%
Dataset size342
Dataset licenceApache-2.0
PublisherGoogle (BIG-bench collaboration); authors at Stanford

What it measures

moral_permissibility gives a short English story, often a trolley-problem variant, and asks whether an action is morally permissible. The target is the majority human answer from the psychology paper that published that story, binarised Yes/No. Authors Allen Nie and Tobias Gerstenberg present it as a diagnostic of whether models track the same factors humans do (personal force, intent, who benefits), not as a test of "correct" ethics. English text.

Task format

Two-option multiple_choice_grade (Yes/No). append_choices_to_input false. Zero-shot in the task keywords. Canary GUID embedded. 342 dummy-model queries. Each example carries a comment with the source paper and, when the authors had it, raw human agreement.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub