Preprint
Jul 2026
DrawingVQA: A Real-World Benchmark for Multi-Depth Visual-Textual Reasoning on Construction Drawings
DrawingVQA, the first benchmark designed to evaluate multimodal large language models (MLLMs) on real-world construction drawings, lays a foundation for domain-specialized multimodal reasoning to allow for advancement on integration of AI-driven understanding and real-world engineering workflows
Yoon Wha Jung, Junryu Fu, M. Golparvar-Fard
· 1 citation