Chart Question Answering with Visual and Logical Reasoning

Prince, Enamul HoqueMasry, Ahmed2022-12-142022-12-142022-06-232022-12-14http://hdl.handle.net/10315/40644Charts are very popular for analyzing data. When exploring charts, people often ask complex reasoning questions that involve several logical and arithmetic operations. They also commonly refer to visual features of a chart in their questions. However, most existing datasets do not focus on such complex reasoning questions as their questions are template-based and answers come from a fixed-vocabulary. In this thesis work, we present a large-scale benchmark covering 9.6K human-written questions and 23.1K questions generated from human-written chart summaries. To address the unique challenges in our benchmark involving visual and logical reasoning, we present transformer-based models that combine visual features and the data table of the chart. Moreover, we propose chart-specific pretraining tasks that improve the visual and logical reasoning skills of our models. While our models achieve the state-of-the-art results on the previous datasets and our benchmark, the evaluation also reveals several challenges in answering complex reasoning questions.Author owns copyright, except where explicitly noted. Please contact the author directly with licensing requests.Computer scienceChart Question Answering with Visual and Logical ReasoningElectronic Thesis or Dissertation2022-12-14ChartsChartQAQA