VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building
Summary: VOCALExplore: pay-as-you-go interactive system for building domain-specific video models that adaptively selects samples by exploiting observed label skew and frames representation/feature choice as a rising-bandit problem. Delivers near-optimal model quality with low visible latency (~1s/iteration) and no expensive preprocessing. (summarized by gpt-5-mini on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Maureen Daum (University of Washington)
- 2. Enhao Zhang (University of Washington)
- 3. Dong He (University of Washington)
- 4. Stephen Mussmann (University of Washington)
- 5. Brandon Haynes (Microsoft)
- 6. Ranjay Krishna (University of Washington)
- 7. Magdalena Balazinska (University of Washington)
BibTeX Citation
@article{daum_vldb23,
title = {{VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building}},
author = {Daum, Maureen and Zhang, Enhao and He, Dong and Mussmann, Stephen and Haynes, Brandon and Krishna, Ranjay and Balazinska, Magdalena},
journal = {PVLDB},
series = {{VLDB} '23},
volume = {16},
number = {13},
pages = {4188--4201},
doi = {10.14778/3625054.3625057},
url = {https://doi.org/10.14778/3625054.3625057},
year = {2023}
}
Incoming Citations (Sorted by Pagerank)
Showing 4 of 4 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 7,447 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB | 5.5240718e-05 |
| 9,645 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD | 5.1453267e-05 |
| 11,310 | Deja Vu: Efficient Video-Language Query Engine with Learning-based Inter-Frame Computation Reuse | 2025 | VLDB | 4.9793485e-05 |
| 11,793 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB | 4.9793485e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 9 of 9 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 281 | Accelerating Machine Learning Inference with Probabilistic Predicates | 2018 | SIGMOD | 0.00022295232 |
| 541 | BlazeIt: Optimizing Declarative Aggregation and Limit Queries for Neural Network-Based Video Analytics | 2020 | VLDB | 0.00016657685 |
| 1,024 | MIRIS: Fast Object Track Queries in Video | 2020 | SIGMOD | 0.00012430491 |
| 2,710 | Panorama: A Data System for Unbounded Vocabulary Querying over Video | 2020 | VLDB | 8.1041775e-05 |
| 2,787 | EVA: A Symbolic Approach to Accelerating Exploratory Video Analytics with Materialized Views | 2022 | SIGMOD | 8.0158999e-05 |
| 3,778 | Optimizing Video Analytics with Declarative Model Relationships | 2023 | VLDB | 7.0259119e-05 |
| 3,884 | FiGO: Fine-Grained Query Optimization in Video Analytics | 2022 | SIGMOD | 6.948464e-05 |
| 5,970 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR | 5.9316412e-05 |
| 9,036 | LANCET: Labeling Complex Data at Scale | 2021 | VLDB | 5.2332952e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 9,541 | DoveDB: A Declarative and Low-Latency Video Database | 2023 | VLDB |
| 2 | 2,710 | Panorama: A Data System for Unbounded Vocabulary Querying over Video | 2020 | VLDB |
| 3 | 11,793 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB |
| 4 | 271 | NoScope: Optimizing Neural Network Queries over Video at Scale | 2017 | VLDB |
| 5 | 3,429 | SVQ: Streaming Video Queries | 2019 | SIGMOD |
| 6 | 9,589 | SketchQL: Video Moment Querying with a Visual Query Interface | 2024 | SIGMOD |
| 7 | 7,447 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB |
| 8 | 9,645 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD |
| 9 | 3,806 | Vaas: Video Analytics At Scale | 2020 | VLDB |
| 10 | 5,970 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR |