DBScholar

Back to papers

Automatic Web-Scale Information Extraction

Summary: Yahoo! demo of web-scale information extraction. Given new websites with semi-structured data mapped to predefined schemas, automatically populate schema objects by extracting values at scale, demonstrating end-to-end, schema-driven extraction robust to site variability and across domains. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
4625
Venue
SIGMOD
Year
2012
Pagerank
5.4574671e-05
Overall Rank
8,246 | 43.43%
DOI
10.1145/2213836.2213912

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{bohannon_sigmod12,
        title = {{Automatic Web-Scale Information Extraction}},
        author = {Bohannon, Philip and Dalvi, Nilesh and Filmus, Yuval and Jacoby, Nori and Keerthi, Sathiya and Kirpal, Alok},
        series = {{SIGMOD} '12},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/2213836.2213912},
        url = {https://dl.acm.org/doi/10.1145/2213836.2213912},
        year = {2012}
}

Incoming Citations (Sorted by Pagerank)

Showing 2 of 2 citing papers.

Rank Citing Paper Year Venue Pagerank
10,414 Visual Template Inference for Data Extraction from Documents 2026 SIGMOD 5.093636e-05
12,242 Knowledge Harvesting in the Big-Data Era 2013 SIGMOD 5.093636e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 6 of 6 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers