VOCAL: Video Organization and Interactive Compositional AnaLytics
Summary: VOCAL: a VDBMS for interactive organization, exploration, and cleaning of heterogeneous video corpora without requiring clean data or pretrained detectors. Supports efficient compositional spatiotemporal queries over multi-object events via optimizations that reduce manual labeling and query costs. (summarized by gpt-5-mini on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Maureen Daum (University of Washington)
- 2. Enhao Zhang (University of Washington)
- 3. Dong He (University of Washington)
- 4. Magdalena Balazinska (University of Washington)
- 5. Brandon Haynes (Microsoft)
- 6. Ranjay Krishna (University of Washington)
- 7. Apryle Craig (University of Washington)
- 8. Aaron Wirsing (University of Washington)
BibTeX Citation
@inproceedings{daum_cidr22,
address = {Amsterdam, Netherlands},
series = {{CIDR} '22},
title = {{VOCAL: Video Organization and Interactive Compositional AnaLytics}},
booktitle = {Proceedings of the {Conference} on {Innovative} {Data} {Systems} {Research}},
author = {Daum, Maureen and Zhang, Enhao and He, Dong and Balazinska, Magdalena and Haynes, Brandon and Krishna, Ranjay and Craig, Apryle and Wirsing, Aaron},
year = {2022}
}
Incoming Citations (Sorted by Pagerank)
Showing 8 of 8 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 2,295 | ThalamusDB: Approximate Query Processing on Multi-Modal Data | 2024 | SIGMOD | 8.6822982e-05 |
| 3,780 | Optimizing Video Analytics with Declarative Model Relationships | 2023 | VLDB | 7.0225859e-05 |
| 7,451 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB | 5.5214568e-05 |
| 9,631 | Interactive Demonstration of EVA | 2023 | VLDB | 5.1468222e-05 |
| 9,653 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD | 5.142891e-05 |
| 10,112 | VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building | 2023 | VLDB | 5.0765311e-05 |
| 11,219 | Scalable Complex Event Processing on Video Streams | 2025 | SIGMOD | 4.9769913e-05 |
| 11,799 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB | 4.9769913e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 14 of 14 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 9,550 | DoveDB: A Declarative and Low-Latency Video Database | 2023 | VLDB |
| 2 | 3,807 | Vaas: Video Analytics At Scale | 2020 | VLDB |
| 3 | 2,787 | EVA: A Symbolic Approach to Accelerating Exploratory Video Analytics with Materialized Views | 2022 | SIGMOD |
| 4 | 5,012 | VisualWorldDB: A DBMS for the Visual World | 2020 | CIDR |
| 5 | 11,602 | Optimizing Video Queries with Declarative Clues | 2024 | VLDB |
| 6 | 9,597 | SketchQL: Video Moment Querying with a Visual Query Interface | 2024 | SIGMOD |
| 7 | 11,799 | EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions | 2023 | VLDB |
| 8 | 10,112 | VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building | 2023 | VLDB |
| 9 | 7,451 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB |
| 10 | 9,653 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD |