Demonstrating ASET: Ad-hoc Structured Exploration of Text Collections
Summary: ASET enables ad-hoc text-to-table exploration in two phases: extract nuggets, then embed and map to a user-defined schema. GUI enables non-experts, no curated pipelines or training data, and yields quick qualitative assessments on unseen corpora. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
Authors
- 1. Benjamin Hättasch (Technical University of Darmstadt)
- 2. Jan-Micha Bodensohn (Technical University of Darmstadt)
- 3. Carsten Binnig (Technical University of Darmstadt)
BibTeX Citation
@inproceedings{hattasch_sigmod22,
title = {{Demonstrating ASET: Ad-hoc Structured Exploration of Text Collections}},
author = {Hättasch, Benjamin and Bodensohn, Jan-Micha and Binnig, Carsten},
series = {{SIGMOD} '22},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/3514221.3520174},
url = {https://dl.acm.org/doi/10.1145/3514221.3520174},
year = {2022}
}
Incoming Citations (Sorted by Pagerank)
Showing 0 of 0 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 1 of 1 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 2,917 | Query-Driven On-The-Fly Knowledge Base Construction | 2018 | VLDB | 7.9661816e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 9,307 | Unify: A System For Unstructured Data Analytics | 2025 | VLDB |
| 2 | 13,690 | The SystemT IDE: An Integrated Development Environment for Information Extraction Rules | 2011 | SIGMOD |
| 3 | 7,482 | Table Extraction and Understanding for Scientific and Enterprise Applications | 2020 | VLDB |
| 4 | 11,980 | Building Structured Databases of Factual Knowledge from Massive Text Corpora | 2017 | SIGMOD |
| 5 | 11,934 | Sherlock: A System for Interactive Summarization of Large Text Collections | 2018 | VLDB |
| 6 | 13,314 | SemExplorer: A User Interface for Semantic Approach to Customized Dataset Search | 2025 | SIGMOD |
| 7 | 13,337 | Accelerating Tabular Inference: Training Data Generation with TENET | 2025 | VLDB |
| 8 | 11,186 | Unstructured Data Fusion for Schema and Data Extraction | 2024 | SIGMOD |
| 9 | 9,294 | TextCube: Automated Construction and Multidimensional Exploration | 2019 | VLDB |
| 10 | 4,992 | Scalable Ad-hoc Entity Extraction from Text Collections | 2008 | VLDB |