EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions
Summary: EQUI-VOCAL synthesizes compositional, declarative video queries via few-shot user inputs (few positives/negatives) plus selective sampling of additional labels, enabling retrieval of complex events without manual query construction. Demo provides a GUI for hands-on query synthesis and exploration of hyperparameter and label-noise effects on retrieval performance. (summarized by gpt-5-mini on Feb 09 2026)
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
Authors
- 1. Enhao Zhang (University of Washington)
- 2. Maureen Daum (University of Washington)
- 3. Dong He (University of Washington)
- 4. Manasi Ganti (University of Washington)
- 5. Brandon Haynes (Microsoft)
- 6. Ranjay Krishna (University of Washington)
- 7. Magdalena Balazinska (University of Washington)
BibTeX Citation
@article{zhang_vldb23,
title = {{EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions}},
author = {Zhang, Enhao and Daum, Maureen and He, Dong and Ganti, Manasi and Haynes, Brandon and Krishna, Ranjay and Balazinska, Magdalena},
journal = {PVLDB},
series = {{VLDB} '23},
volume = {16},
number = {12},
pages = {3978--3981},
doi = {10.14778/3611540.3611600},
url = {https://doi.org/10.14778/3611540.3611600},
year = {2023}
}
Incoming Citations (Sorted by Pagerank)
Showing 1 of 1 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 9,464 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD | 5.2634238e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 6 of 6 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 295 | Accelerating Machine Learning Inference with Probabilistic Predicates | 2018 | SIGMOD | 0.00022238183 |
| 3,365 | Spatial and Temporal Constrained Ranked Retrieval over Videos | 2022 | VLDB | 7.4763252e-05 |
| 4,213 | SVQ++: Querying for Object Interactions in Video Streams | 2020 | SIGMOD | 6.8279962e-05 |
| 5,966 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR | 6.0255527e-05 |
| 8,362 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB | 5.4432545e-05 |
| 9,926 | VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building | 2023 | VLDB | 5.1955087e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 11,268 | Optimizing Video Queries with Declarative Clues | 2024 | VLDB |
| 2 | 3,410 | SVQ: Streaming Video Queries | 2019 | SIGMOD |
| 3 | 11,543 | VoiceQuerySystem: A Voice-driven Database Querying System Using Natural Language Questions | 2022 | SIGMOD |
| 4 | 9,421 | SketchQL Demonstration: Zero-shot Video Moment Querying with Sketches | 2024 | VLDB |
| 5 | 9,408 | SketchQL: Video Moment Querying with a Visual Query Interface | 2024 | SIGMOD |
| 6 | 9,464 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD |
| 7 | 5,966 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR |
| 8 | 4,593 | Demonstration of SpeakQL: Speech-driven Multimodal Querying of Structured Data | 2019 | SIGMOD |
| 9 | 1,828 | Making the Case for Query-by-Voice with EchoQuery | 2016 | SIGMOD |
| 10 | 8,362 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB |