Semantic-equivalence-aware evaluation for document parsing
Develop evaluation methodologies for document parsing that are semantic-equivalence-aware, explicitly accounting for both format-level ambiguity—such as HTML versus Markdown representations for tables and alternative LaTeX command choices that encode the same mathematical content—and structural-level ambiguity—such as representing a bilingual aligned word list either as line-by-line paired text blocks or as a two-column table—so that different but semantically equivalent outputs receive fair and consistent scores.
References
Developing semantic-equivalence-aware evaluation methods that account for both format and structural ambiguity remains an open problem.
Three entries have no mechanical answer and are listed as open, which we regard as an honest result rather than an omission.