VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building
Summary: VOCALExplore: pay-as-you-go interactive system for building domain-specific video models that adaptively selects samples by exploiting observed label skew and frames representation/feature choice as a rising-bandit problem. Delivers near-optimal model quality with low visible latency (~1s/iteration) and no expensive preprocessing. (summarized by gpt-5-mini on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Maureen Daum (University of Washington)
- 2. Enhao Zhang (University of Washington)
- 3. Dong He (University of Washington)
- 4. Stephen Mussmann (University of Washington)
- 5. Brandon Haynes (Microsoft)
- 6. Ranjay Krishna (University of Washington)
- 7. Magdalena Balazinska (University of Washington)
BibTeX Citation
@article{daum_vldb23,
title = {{VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building}},
author = {Daum, Maureen and Zhang, Enhao and He, Dong and Mussmann, Stephen and Haynes, Brandon and Krishna, Ranjay and Balazinska, Magdalena},
journal = {PVLDB},
series = {{VLDB} '23},
volume = {16},
number = {13},
pages = {4188--4201},
doi = {10.14778/3625054.3625057},
url = {https://doi.org/10.14778/3625054.3625057},
year = {2023}
}
Incoming Citations (Sorted by Pagerank)
Showing 4 of 4 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 7,451 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB | 5.5214568e-05 |
| 9,653 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD | 5.142891e-05 |
| 11,318 | Deja Vu: Efficient Video-Language Query Engine with Learning-based Inter-Frame Computation Reuse | 2025 | VLDB | 4.9769913e-05 |
| 11,799 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB | 4.9769913e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 9 of 9 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 282 | Accelerating Machine Learning Inference with Probabilistic Predicates | 2018 | SIGMOD | 0.00022302793 |
| 541 | BlazeIt: Optimizing Declarative Aggregation and Limit Queries for Neural Network-Based Video Analytics | 2020 | VLDB | 0.00016663833 |
| 1,023 | MIRIS: Fast Object Track Queries in Video | 2020 | SIGMOD | 0.00012430702 |
| 2,705 | Panorama: A Data System for Unbounded Vocabulary Querying over Video | 2020 | VLDB | 8.1096249e-05 |
| 2,787 | EVA: A Symbolic Approach to Accelerating Exploratory Video Analytics with Materialized Views | 2022 | SIGMOD | 8.0121053e-05 |
| 3,780 | Optimizing Video Analytics with Declarative Model Relationships | 2023 | VLDB | 7.0225859e-05 |
| 3,877 | FiGO: Fine-Grained Query Optimization in Video Analytics | 2022 | SIGMOD | 6.9496983e-05 |
| 5,948 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR | 5.934521e-05 |
| 9,044 | LANCET: Labeling Complex Data at Scale | 2021 | VLDB | 5.2308178e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 9,550 | DoveDB: A Declarative and Low-Latency Video Database | 2023 | VLDB |
| 2 | 2,705 | Panorama: A Data System for Unbounded Vocabulary Querying over Video | 2020 | VLDB |
| 3 | 11,799 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB |
| 4 | 271 | NoScope: Optimizing Neural Network Queries over Video at Scale | 2017 | VLDB |
| 5 | 3,429 | SVQ: Streaming Video Queries | 2019 | SIGMOD |
| 6 | 9,597 | SketchQL: Video Moment Querying with a Visual Query Interface | 2024 | SIGMOD |
| 7 | 7,451 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB |
| 8 | 9,653 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD |
| 9 | 3,807 | Vaas: Video Analytics At Scale | 2020 | VLDB |
| 10 | 5,948 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR |