arxiv
PublishedSeptember 3, 2026 at 4:00 AM
—neutral
DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents
Publisher summary· verbatim
arXiv:2609.02059v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved strong performance on structured visual understanding tasks such as chart and document question answering. However, existing benchmarks typically evaluate these domains in isolation, leaving undere
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivMeta-ethics and AI: exploring the novel meta-ethical questions in the era of AI2harxivSSAKG 2.0: An Open-Source Package for Structural Associative Sequence Memory and Context-Based Retrieval2harxivEpistemic Sybil Resistance: Multiplying AI Agents Without Multiplying Evidence2harxivMASkills: Continual Skills Optimization for Multi-Agent LLM Systems2hThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗