VOCAL: Video Organization and Interactive Compositional AnaLytics
Summary: VOCAL: a VDBMS for interactive organization, exploration, and cleaning of heterogeneous video corpora without requiring clean data or pretrained detectors. Supports efficient compositional spatiotemporal queries over multi-object events via optimizations that reduce manual labeling and query costs. (summarized by gpt-5-mini on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Maureen Daum (University of Washington)
- 2. Enhao Zhang (University of Washington)
- 3. Dong He (University of Washington)
- 4. Magdalena Balazinska (University of Washington)
- 5. Brandon Haynes (Microsoft)
- 6. Ranjay Krishna (University of Washington)
- 7. Apryle Craig (University of Washington)
- 8. Aaron Wirsing (University of Washington)
BibTeX Citation
@inproceedings{daum_cidr22,
address = {Amsterdam, Netherlands},
series = {{CIDR} '22},
title = {{VOCAL: Video Organization and Interactive Compositional AnaLytics}},
booktitle = {Proceedings of the {Conference} on {Innovative} {Data} {Systems} {Research}},
author = {Daum, Maureen and Zhang, Enhao and He, Dong and Balazinska, Magdalena and Haynes, Brandon and Krishna, Ranjay and Craig, Apryle and Wirsing, Aaron},
year = {2022}
}
Incoming Citations (Sorted by Pagerank)
Showing 8 of 8 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 2,450 | ThalamusDB: Approximate Query Processing on Multi-Modal Data | 2024 | SIGMOD | 8.4474092e-05 |
| 3,778 | Optimizing Video Analytics with Declarative Model Relationships | 2023 | VLDB | 7.0259119e-05 |
| 7,447 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB | 5.5240718e-05 |
| 9,624 | Interactive Demonstration of EVA | 2023 | VLDB | 5.1492598e-05 |
| 9,645 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD | 5.1453267e-05 |
| 10,108 | VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building | 2023 | VLDB | 5.0789354e-05 |
| 11,210 | Scalable Complex Event Processing on Video Streams | 2025 | SIGMOD | 4.9793485e-05 |
| 11,793 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB | 4.9793485e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 14 of 14 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 9,541 | DoveDB: A Declarative and Low-Latency Video Database | 2023 | VLDB |
| 2 | 3,806 | Vaas: Video Analytics At Scale | 2020 | VLDB |
| 3 | 2,787 | EVA: A Symbolic Approach to Accelerating Exploratory Video Analytics with Materialized Views | 2022 | SIGMOD |
| 4 | 5,026 | VisualWorldDB: A DBMS for the Visual World | 2020 | CIDR |
| 5 | 11,596 | Optimizing Video Queries with Declarative Clues | 2024 | VLDB |
| 6 | 9,589 | SketchQL: Video Moment Querying with a Visual Query Interface | 2024 | SIGMOD |
| 7 | 11,793 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB |
| 8 | 10,108 | VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building | 2023 | VLDB |
| 9 | 7,447 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB |
| 10 | 9,645 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD |