DBScholar

Back to papers

Interesting-Phrase Mining for Ad-Hoc Text Analytics

Summary: Introduces a phrase-centric framework for ad-hoc text analytics, prioritizing multi-word phrases that are frequent in a subset yet rare in the full corpus. Develops preprocessing, indexing, and top-k search methods for scalable discovery, validated on a large NYT corpus. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h90171351c410357d
Venue
VLDB
Year
2010
Pagerank
5.5164354e-05
Overall Rank
7,475 | 49.75%
DOI
10.14778/1920841.1921007

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{bedathur_vldb10,
        title = {{Interesting-Phrase Mining for Ad-Hoc Text Analytics}},
        author = {Bedathur, Srikanta and Berberich, Klaus and Dittrich, Jens and Mamoulis, Nikos and Weikum, Gerhard},
        journal = {PVLDB},
        series = {{VLDB} '10},
        volume = {3},
        number = {1},
        pages = {1348--1359},
        doi = {10.14778/1920841.1921007},
        url = {https://doi.org/10.14778/1920841.1921007},
        year = {2010}
}

Incoming Citations (Sorted by Pagerank)

Showing 1 of 1 citing papers.

Rank Citing Paper Year Venue Pagerank
8,239 Mining Quality Phrases from Massive Text Corpora 2015 SIGMOD 5.3708698e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 5 of 5 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
164 Mining Frequent Patterns without Candidate Generation 2000 SIGMOD 0.00027412227
2,760 BlogScope: A System for Online Analysis of High Volume Text Streams 2007 VLDB 8.0493804e-05
3,536 Multidimensional Content eXploration 2008 VLDB 7.2241245e-05
4,361 Multi-Structural Databases 2005 PODS 6.6385952e-05
6,577 Efficient Implementation of Large-Scale Multi-Structural Databases 2005 VLDB 5.7447944e-05
Previous Page 1 / 1 Next

Semantically Similar Papers