Facilitating Document Reading by Linking Text and Tables

Kim, Dae Hyun; Hoque, Enamul; Kim, Juho; Agrawala, Maneesh

doi:10.1145/3242587.3242617

Citation Details

Facilitating Document Reading by Linking Text and Tables

Document authors commonly use tables to support arguments presented in the text. But, because tables are usually separate from the main body text, readers must split their attention between different parts of the document. We present an interactive document reader that automatically links document text with corresponding table cells. Readers can select a sentence (or tables cells) and our reader highlights the relevant table cells (or sentences). We provide an automatic pipeline for extracting such references between sentence text and table cells for existing PDF documents that combines structural analysis of tables with natural language processing and rule-based matching. On a test corpus of 330 (sentence, table) pairs, our pipeline correctly extracts 48.8% of the references. An additional 30.5% contain only false negatives (FN) errors -- the reference is missing table cells. The remaining 20.7% contain false positives (FP) errors -- the reference includes extraneous table cells and could therefore mislead readers. A user study finds that despite such errors, our interactive document reader helps readers match sentences with corresponding table cells more accurately and quickly than a baseline document reader. more »

Award ID(s):: 1714647

PAR ID:: 10098281

Author(s) / Creator(s):: Kim, Dae Hyun; Hoque, Enamul; Kim, Juho; Agrawala, Maneesh

Date Published:: 2018-01-01

Journal Name:: Proceedings of the 31st Annual ACM Symposium on User Interface Software and Technology

Page Range / eLocation ID:: 423 to 434

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
https://doi.org/10.1145/3242587.3242617

More Like this