Preprint
Aug 2026
LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning
This work evaluates 10 state-of-the-art MLLMs and examines three factors that influence performance: reasoning patterns, auxiliary tools, and robustness to image perturbations, showing that MLLM accuracy decreases and varies substantially as computational complexity increases.
Ziyan Xiao, Yinghao Zhu, Wenting Zhang et al.
· 0 citations