Advancing Chart Question Answering with Robust Chart Component Recognition

Zheng, Hanwen; Wang, Sijia; Thomas, Chris; Huang, Lifu

doi:10.1109/WACV61041.2025.00560

Citation Details

This content will become publicly available on February 26, 2026

Advancing Chart Question Answering with Robust Chart Component Recognition

Chart comprehension presents significant challenges for machine learning models due to the diverse and intricate shapes of charts. Existing multimodal methods often over-look these visual features or fail to integrate them effectively for Chart Question Answering. To address this, we introduce CHARTFORMER, a unified framework that enhances chart component recognition by accurately identifying and classifying components such as bars, lines, pies, titles, legends, and axes. Additionally, we propose a novel Question-guided Deformable Co-Attention (QDCAt) mechanism, which fuses chart features encoded by Chart-former with the given question, leveraging the question's guidance to ground the correct answer. Extensive experiments demonstrate a 3.2% improvement in mAP over the baselines for chart component recognition. For ChartQA and OpenCQA tasks, our approach achieves improvements of 15.4% in accuracy and 0.8 in BLEU score, respectively, underscoring the robustness of our solution for detailed visual data interpretation across various applications. more »

Award ID(s):: 2238940

PAR ID:: 10629367

Author(s) / Creator(s):: Zheng, Hanwen; Wang, Sijia; Thomas, Chris; Huang, Lifu

Publisher / Repository:: IEEE

Date Published:: 2025-02-26

ISBN:: 979-8-3315-1083-1

Page Range / eLocation ID:: 5741 to 5750

Format(s):: Medium: X

Location:: Tucson, AZ, USA

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
This content will become publicly available on February 26, 2026
Conference Paper:
https://doi.org/10.1109/WACV61041.2025.00560

More Like this