Visual Segmentation for Information Extraction from Heterogeneous Visually Rich Documents
Summary: VS2 segments visually rich documents into logical blocks via document-type-agnostic cues. A distantly supervised search-and-select uses block boundaries to locate entities, outperforming text-only IE across three heterogeneous datasets. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Ritesh Sarkhel (Ohio State University)
- 2. Arnab Nandi (Ohio State University)
BibTeX Citation
@inproceedings{sarkhel_sigmod19,
title = {{Visual Segmentation for Information Extraction from Heterogeneous Visually Rich Documents}},
author = {Sarkhel, Ritesh and Nandi, Arnab},
series = {{SIGMOD} '19},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/3299869.3319867},
url = {https://dl.acm.org/doi/10.1145/3299869.3319867},
year = {2019}
}
Incoming Citations (Sorted by Pagerank)
Showing 4 of 4 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 8,176 | Improving Information Extraction from Visually Rich Documents using Visual Span Representations | 2021 | VLDB | 5.3835163e-05 |
| 9,581 | Glean: Structured Extractions from Templatic Documents | 2021 | VLDB | 5.1571823e-05 |
| 10,800 | MGRAG: Semantic Subgraph Matching and Graph-Aware Caching for Multimodal Retrieval-Augmented Generation | 2026 | VLDB | 4.9793485e-05 |
| 11,767 | Self-Training for Label-Efficient Information Extraction from Semi-Structured Web-Pages | 2023 | VLDB | 4.9793485e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 4 of 4 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 515 | RoadRunner: Towards Automatic Data Extraction from Large Web Sites | 2001 | VLDB | 0.00016970049 |
| 3,061 | Fonduer: Knowledge Base Construction from Richly Formatted Data | 2018 | SIGMOD | 7.6927483e-05 |
| 4,111 | Enterprise Information Extraction: Recent Developments and Open Challenges | 2010 | SIGMOD | 6.8006098e-05 |
| 6,292 | Extracting Logical Hierarchical Structure of HTML Documents Based on Headings | 2015 | VLDB | 5.8215681e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 11,764 | Automatic Road Extraction with Multi-Source Data Revisited: Completeness, Smoothness and Discrimination | 2023 | VLDB |
| 2 | 3,588 | Using the Structure of Web Sites for Automatic Segmentation of Tables | 2004 | SIGMOD |
| 3 | 4,495 | Evaluating Temporal Queries Over Video Feeds | 2021 | SIGMOD |
| 4 | 6,292 | Extracting Logical Hierarchical Structure of HTML Documents Based on Headings | 2015 | VLDB |
| 5 | 4,736 | Scalable Ad-hoc Entity Extraction from Text Collections | 2008 | VLDB |
| 6 | 3,411 | Spatial and Temporal Constrained Ranked Retrieval over Videos | 2022 | VLDB |
| 7 | 11,692 | Unsupervised Hashing with Semantic Concept Mining | 2023 | SIGMOD |
| 8 | 3,512 | Efficient Approximate Entity Extraction with Edit Distance Constraints | 2009 | SIGMOD |
| 9 | 10,605 | Visual Template Inference for Data Extraction from Documents | 2026 | SIGMOD |
| 10 | 8,176 | Improving Information Extraction from Visually Rich Documents using Visual Span Representations | 2021 | VLDB |