CLIcK

Korean multiple-choice exam of cultural and linguistic knowledge: 1,995 questions in eleven categories, drawn from official exams and textbooks.

Also known as: CLIcK: Cultural and Linguistic Intelligence in Korean, Cultural and Linguistic Intelligence in Korean

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categoryknowledge
SubcategoryKorean cultural and linguistic multiple-choice QA from exams and textbooks
Page statusactive
Metricaccuracy and length-normalized accuracy (acc, acc_norm)
Directionhigher_is_better
Unit%
Dataset size1995
PublisherKAIST (School of Computing and GSAI)

What it measures

click is EleutherAI lm-evaluation-harness's group for CLIcK (Cultural and Linguistic Intelligence in Korean). The model reads a Korean question, often with a short context, and picks a lettered choice. Items test Korean language (textual, grammatical, functional knowledge) and Korean culture (society, tradition, politics, economy, law, history, geography, popular culture). Sources are official exams plus GPT-4 questions written from the KIIP textbook and then validated. This is a Korean-centric knowledge test, not a translated English quiz and not [korbench](korbench.md) (which is English knowledge-orthogonal reasoning despite the similar id).

Task format

Multiple choice in Korean. lm-eval output_type is multiple_choice. The Korean prompt in utils.get_context always lists A–D. Scoring helpers get_choices and get_target use A–E when the example id contains CSAT, else A–D. The Hugging Face split used as both test_split and fewshot_split is named train. Groups: click (all 11), click_lang (3), click_cul (8).

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub