DBScholar

Back to papers

Optimal Algorithms for Crawling a Hidden Database in the Web

Summary: Algorithms to extract all tuples from a hidden web database via a query-only interface, even when results are partial. Provably efficient in the worst case and asymptotically optimal, with extensive experiments on real datasets. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
10538
Venue
VLDB
Year
2012
Pagerank
5.2352357e-05
Overall Rank
9,683 | 33.57%
DOI
10.14778/2350229.2350240

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{sheng_vldb12,
        title = {{Optimal Algorithms for Crawling a Hidden Database in the Web}},
        author = {Sheng, Cheng and Zhang, Nan and Tao, Yufei and Jin, Xin},
        journal = {PVLDB},
        series = {{VLDB} '12},
        volume = {5},
        number = {11},
        pages = {1112--1123},
        doi = {10.14778/2350229.2350240},
        url = {https://doi.org/10.14778/2350229.2350240},
        year = {2012}
}

Incoming Citations (Sorted by Pagerank)

Showing 6 of 6 citing papers.

Rank Citing Paper Year Venue Pagerank
8,378 Discovering the Skyline of Web Databases 2016 VLDB 5.4387677e-05
8,705 Progressive Deep Web Crawling Through Keyword Queries For Data Enrichment 2019 SIGMOD 5.3807913e-05
9,650 Aggregate Estimation Over Dynamic Hidden Web Databases 2014 VLDB 5.2427667e-05
12,083 Query Reranking As A Service 2016 VLDB 5.093636e-05
12,285 Rank Discovery From Web Databases 2013 VLDB 5.093636e-05
13,612 HDBTracker: Monitoring the Aggregates On Dynamic Hidden Web Databases 2014 VLDB -
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 10 of 10 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers