EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions
Summary: EQUI-VOCAL synthesizes compositional, declarative video queries via few-shot user inputs (few positives/negatives) plus selective sampling of additional labels, enabling retrieval of complex events without manual query construction. Demo provides a GUI for hands-on query synthesis and exploration of hyperparameter and label-noise effects on retrieval performance. (summarized by gpt-5-mini on Feb 09 2026)
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
Authors
- 1. Enhao Zhang (University of Washington)
- 2. Maureen Daum (University of Washington)
- 3. Dong He (University of Washington)
- 4. Manasi Ganti (University of Washington)
- 5. Brandon Haynes (Microsoft)
- 6. Ranjay Krishna (University of Washington)
- 7. Magdalena Balazinska (University of Washington)
BibTeX Citation
@article{zhang_vldb23,
title = {{EQUI-VOCAL Demonstration: Synthesizing Video Queries from User Interactions}},
author = {Zhang, Enhao and Daum, Maureen and He, Dong and Ganti, Manasi and Haynes, Brandon and Krishna, Ranjay and Balazinska, Magdalena},
journal = {PVLDB},
series = {{VLDB} '23},
volume = {16},
number = {12},
pages = {3978--3981},
doi = {10.14778/3611540.3611600},
url = {https://doi.org/10.14778/3611540.3611600},
year = {2023}
}
Incoming Citations (Sorted by Pagerank)
Showing 1 of 1 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 9,645 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD | 5.1453267e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 6 of 6 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 281 | Accelerating Machine Learning Inference with Probabilistic Predicates | 2018 | SIGMOD | 0.00022295232 |
| 3,411 | Spatial and Temporal Constrained Ranked Retrieval over Videos | 2022 | VLDB | 7.3253532e-05 |
| 4,286 | SVQ++: Querying for Object Interactions in Video Streams | 2020 | SIGMOD | 6.6873151e-05 |
| 5,970 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR | 5.9316412e-05 |
| 7,447 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB | 5.5240718e-05 |
| 10,108 | VOCALExplore: Pay-as-You-Go Video Data Exploration and Model Building | 2023 | VLDB | 5.0789354e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 11,596 | Optimizing Video Queries with Declarative Clues | 2024 | VLDB |
| 2 | 3,429 | SVQ: Streaming Video Queries | 2019 | SIGMOD |
| 3 | 11,852 | VoiceQuerySystem: A Voice-driven Database Querying System Using Natural Language Questions | 2022 | SIGMOD |
| 4 | 9,602 | SketchQL Demonstration: Zero-shot Video Moment Querying with Sketches | 2024 | VLDB |
| 5 | 9,589 | SketchQL: Video Moment Querying with a Visual Query Interface | 2024 | SIGMOD |
| 6 | 9,645 | Self-Enhancing Video Data Management System for Compositional Events with Large Language Models | 2025 | SIGMOD |
| 7 | 5,970 | VOCAL: Video Organization and Interactive Compositional AnaLytics | 2022 | CIDR |
| 8 | 4,694 | Demonstration of SpeakQL: Speech-driven Multimodal Querying of Structured Data | 2019 | SIGMOD |
| 9 | 1,877 | Making the Case for Query-by-Voice with EchoQuery | 2016 | SIGMOD |
| 10 | 7,447 | EQUI-VOCAL: Synthesizing Queries for Compositional Video Events from Limited User Interactions | 2023 | VLDB |