Dr. DocBench

Dr. DocBench evaluates expert-level document parsing on difficult multilingual pages with layout, reading-order, and domain-specific annotations.

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categorymultimodal
Subcategorydocument parsing
Metricparsing quality
Directionhigher_is_better
Unitscore
Dataset size4514
PublisherDr. DocBench authors

What it measures

Dr. DocBench tests vision-language models and document parsers on challenging pages from long multilingual books. It spans 52 BISAC subject domains and targets chemical formulae, music notation, complex tables, cross-page layouts, and other structures where modern parsers fail.

Task format

Document-page parsing with page- and block-level structural annotations.

Models reporting this benchmark

No model card in ModelSpec reports this benchmark yet.

Data

This page as JSON · Edit on GitHub