VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building
Summary: VOCALExplore: pay-as-you-go interactive system for building domain-specific video models that adaptively selects samples by exploiting observed label skew and frames representation/feature choice as a rising-bandit problem. Delivers near-optimal model quality with low visible latency (~1s/iteration) and no expensive preprocessing. (summarized by gpt-5-mini on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Maureen Daum (University of Washington)
- 2. Enhao Zhang (University of Washington)
- 3. Dong He (University of Washington)
- 4. Stephen Mussmann (University of Washington)
- 5. Brandon Haynes (Microsoft)
- 6. Ranjay Krishna (University of Washington)
- 7. Magdalena Balazinska (University of Washington)
BibTeX Citation
@article{daum_vldb23,
title = {{VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building}},
author = {Daum, Maureen and Zhang, Enhao and He, Dong and Mussmann, Stephen and Haynes, Brandon and Krishna, Ranjay and Balazinska, Magdalena},
journal = {PVLDB},
series = {{VLDB} '23},
volume = {16},
number = {13},
pages = {4188--4201},
doi = {10.14778/3625054.3625057},
url = {https://doi.org/10.14778/3625054.3625057},
year = {2023}
}
Incoming Citations (Sorted by Pagerank)
Showing 4 of 4 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 8,362 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB | 5.4432545e-05 |
| 9,464 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD | 5.2634238e-05 |
| 10,917 | Deja Vu: Efficient Video-Language Query Engine with Learning-based Inter-Frame Computation Reuse | 2025 | VLDB | 5.093636e-05 |
| 11,483 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB | 5.093636e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 9 of 9 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 295 | Accelerating Machine Learning Inference with Probabilistic Predicates | 2018 | SIGMOD | 0.00022238183 |
| 569 | BlazeIt: Optimizing Declarative Aggregation and Limit Queries for Neural Network-Based Video Analytics | 2020 | VLDB | 0.00016348191 |
| 1,042 | MIRIS: Fast Object Track Queries in Video | 2020 | SIGMOD | 0.00012451966 |
| 2,702 | Panorama: A Data System for Unbounded Vocabulary Querying over Video | 2020 | VLDB | 8.2342712e-05 |
| 2,933 | EVA: A Symbolic Approach to Accelerating Exploratory Video Analytics with Materialized Views | 2022 | SIGMOD | 7.9474026e-05 |
| 4,007 | Optimizing Video Analytics with Declarative Model Relationships | 2023 | VLDB | 6.9632395e-05 |
| 4,375 | FiGO: Fine-Grained Query Optimization in Video Analytics | 2022 | SIGMOD | 6.7369552e-05 |
| 5,966 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR | 6.0255527e-05 |
| 8,876 | LANCET: Labeling Complex Data at Scale | 2021 | VLDB | 5.3534114e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 9,925 | DoveDB: A Declarative and Low-Latency Video Database | 2023 | VLDB |
| 2 | 2,702 | Panorama: A Data System for Unbounded Vocabulary Querying over Video | 2020 | VLDB |
| 3 | 11,483 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB |
| 4 | 284 | NoScope: Optimizing Neural Network Queries over Video at Scale | 2017 | VLDB |
| 5 | 3,410 | SVQ: Streaming Video Queries | 2019 | SIGMOD |
| 6 | 9,408 | SketchQL: Video Moment Querying with a Visual Query Interface | 2024 | SIGMOD |
| 7 | 8,362 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB |
| 8 | 9,464 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD |
| 9 | 3,756 | Vaas: Video Analytics At Scale | 2020 | VLDB |
| 10 | 5,966 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR |