AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

ICDAR2026 Competition on Multimodal Reasoning over Documents in Multiple Domains

arXiv · AI, language, vision and robotics · article · Sep 8, 2026 · UTC

In this report we present results of the ICDAR2026 Competition on Multimodal Reasoning over Documents in Multiple Domains. This competition aimed to advance research in document understanding through the task of Visual Question Answering (VQA). Building upon previous DocVQA benchmarks, this competition introduces challenging reasoning questions over a diverse collection of documents spanning eight domains, including business reports, scientific papers, slides, posters, maps, comics, infographics, and engineering drawings. The competition concluded with 20 valid submissions from 8 teams spannin

Read original source ↗ Open in workspace

recordType
paper
region
Global

Evidence & attribution

First collected: 2026-09-25T08:12:34.207Z. This is not the publication date.