OpenCompass suite of 21 biology tasks on opencompass/biology-instruction, scored with MCC, correlation, R², AUC, accuracy and EC-number Fmax.
unassessed
| Category | domain |
|---|---|
| Subcategory | DNA, RNA, protein and multi-sequence property prediction |
| Page status | unknown |
| Metric | task-specific (MCC, PCC, Spearman, R², AUC, accuracy, EC Fmax, Mixed) |
| Direction | higher_is_better |
| Publisher | OpenCompass |
OpenCompass biodata is a generation benchmark over biological sequence problems. Each item is a prompt that asks a model to predict a property of DNA, RNA, protein, or a multi-sequence interaction, then to put the answer in \\boxed{}. Tasks include DNA classification (cpd, emp, pd, transcription-factor binding), enhancer activity regression, RNA isoform and ribosome-loading regression, RNA modification labels, protein solubility, fluorescence, stability, thermostability, enzyme commission (EC) numbers, antibody–antigen and RNA–protein interaction, and siRNA efficiency. The intended skill is biological prediction from sequence context, not general reading comprehension.
Zero-shot generation with a biology-expert system prompt. Two OpenCompass configs differ only in prompt template class (PromptTemplate vs RawPromptTemplate). Answers are parsed from \\boxed{} or from a JSON object for dict-valued labels.
No model card in ModelSpec reports this benchmark yet.