Dae Hyun Kim: Facilitating Document Reading by Linking Text and Tables
https://www.youtube.com/watch?v=kvlIepSfwvI
https://gyazo.com/7a1112a25b882043b4b18458c62f0b54
タイトル
ソース
Proceedings of the 31st Annual ACM Symposium on User Interface Software and Technology (UIST2018) ページ
423-434
年
2018
ISBN
978-1-4503-5948-1
著者
概要
Document authors commonly use tables to support arguments presented in the text. But, because tables are usually separate from the main body text, readers must split their attention between different parts of the document. We present an interactive document reader that automatically links document text with corresponding table cells. Readers can select a sentence (or tables cells) and our reader highlights the relevant table cells (or sentences). We provide an automatic pipeline for extracting such references between sentence text and table cells for existing PDF documents that combines structural analysis of tables with natural language processing and rule-based matching. On a test corpus of 330 (sentence, table) pairs, our pipeline correctly extracts 48.8% of the references. An additional 30.5% contain only false negatives (FN) errors -- the reference is missing table cells. The remaining 20.7% contain false positives (FP) errors -- the reference includes extraneous table cells and could therefore mislead readers. A user study finds that despite such errors, our interactive document reader helps readers match sentences with corresponding table cells more accurately and quickly than a baseline document reader.
内容
PDF中の表の中のエントリとテキストの関連を自動的に表示する
テキストの一部を選択すると関連するエントリがハイライトされるなど
正しくリンクされる確率は50%程度らしい
普通は表をチェックしないのだがこのシステム上だと表をちゃんとチェックした人がいたらしい
コメント
増井俊之.icon
そんなに便利なものだろうか?
間違いが多いと使う気にならないかもしれない
特にFalse Positiveが気になるようだ
判定はかなりアドホック