Arabic Finance (HELM Arabic Enterprise)

HELM's Arabic Enterprise finance set: 299 textbook-derived items in three formats (3-way MCQ, yes/no, numeric calculation), in Arabic or English.

Also known as: arabic_finance_mcq, arabic_finance_bool, arabic_finance_calculation, Arabic Enterprise finance

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categorydomain
SubcategoryArabic (and English) finance textbook QA: 3-way MCQ, yes/no, numeric calculation
Page statusunknown
Metricexact_match (MCQ); quasi_exact_match (bool, schema headline); calculation_accuracy (calculation)
Directionhigher_is_better
Unit%
Dataset size299
Dataset licenceCC-BY-4.0
PublisherStanford CRFM (HELM Arabic Enterprise)

What it measures

arabic_finance is HELM's finance slice of the stanford-crfm/arabic-enterprise dataset. Each item is a short finance question drawn, per HELM's schema, from English-language finance textbooks and machine-translated into Arabic. The same 299 rows are stored with English and Arabic question, choice and answer fields. HELM splits them into three task formats: 23 three-option multiple-choice questions (task=mcq), 57 yes/no verifications (task=bool), and 219 numeric calculation problems (task=calcu). A run is one format and one language (default Arabic). It is text-only professional-finance QA, not [financebench](financebench.md) (English SEC-filing QA) and not [financeiq](financeiq.md).

Task format

Three HELM run specs share one CSV. MCQ: pick one of three labelled choices (CSV uses A/B/C; the Arabic adapter remaps prefixes to أ/ب/ج). Bool: answer نعم/لا or Yes/No. Calculation: write reasoning, then a numeric answer inside \\boxed{}; an LLM annotator judges mathematical equivalence.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub