| 5,800 |
AQWA: Adaptive Query-Workload-Aware Partitioning of Big Spatial Data |
2015 |
VLDB |
5.3218628e-05 |
| 5,817 |
BlinkML: Efficient Maximum Likelihood Estimation with Probabilistic Guarantees |
2019 |
SIGMOD |
5.3154329e-05 |
| 5,850 |
HadoopDB in Action: Building Real World Applications |
2010 |
SIGMOD |
5.3008257e-05 |
| 5,908 |
Building Wavelet Histograms on Large Data in MapReduce |
2012 |
VLDB |
5.2731311e-05 |
| 5,985 |
The Era of Big Spatial Data |
2017 |
VLDB |
5.2399365e-05 |
| 6,120 |
REEF: Retainable Evaluator Execution Framework |
2015 |
SIGMOD |
5.199013e-05 |
| 6,135 |
Fast Data in the Era of Big Data: Twitter's Real-Time Related Query Suggestion Architecture |
2013 |
SIGMOD |
5.1908958e-05 |
| 6,366 |
Good to the Last Bit: Data-Driven Encoding with CodecDB |
2021 |
SIGMOD |
5.0892171e-05 |
| 6,403 |
Just-In-Time Data Virtualization: Lightweight Data Management with ViDa |
2015 |
CIDR |
5.0717043e-05 |
| 6,477 |
Towards Unified Ad-hoc Data Processing |
2014 |
SIGMOD |
5.0408007e-05 |
| 6,661 |
Scalable Querying of Nested Data |
2021 |
VLDB |
4.9663934e-05 |
| 6,817 |
Hadoop's Adolescence: An analysis of Hadoop usage in scientific workloads |
2013 |
VLDB |
4.9110239e-05 |
| 6,833 |
An Algebraic Approach for Data-Centric Scientific Workflows |
2011 |
VLDB |
4.907413e-05 |
| 7,079 |
JetScope: Reliable and Interactive Analytics at Cloud Scale |
2015 |
VLDB |
4.8353804e-05 |
| 7,110 |
Wide Table Layout Optimization based on Column Ordering and Duplication |
2017 |
SIGMOD |
4.8228761e-05 |
| 7,195 |
BSMA: A Benchmark for Analytical Queries over Social Media Data |
2014 |
VLDB |
4.799032e-05 |
| 7,205 |
Kodiak: Leveraging Materialized Views For Very Low-Latency Analytics Over High-Dimensional Web-Scale Data |
2016 |
VLDB |
4.7965293e-05 |
| 7,260 |
Online Expansion of Large-scale Data Warehouses |
2011 |
VLDB |
4.781278e-05 |
| 7,269 |
Oracle In-Database Hadoop: When MapReduce Meets RDBMS |
2012 |
SIGMOD |
4.7768528e-05 |
| 7,293 |
Optimization for iterative queries on MapReduce |
2014 |
VLDB |
4.7668182e-05 |
| 7,533 |
Enabling Efficient and General Subpopulation Analytics in Multidimensional Data Streams |
2022 |
VLDB |
4.7134753e-05 |
| 7,824 |
A Survey and Experimental Comparison of Distributed SPARQL Engines for Very Large RDF Data |
2017 |
VLDB |
4.6390831e-05 |
| 7,879 |
Emerging Trends in the Enterprise Data Analytics: Connecting Hadoop and DB2 Warehouse |
2011 |
SIGMOD |
4.6253726e-05 |
| 7,899 |
Building Highly-Optimized, Low-Latency Pipelines for Genomic Data Analysis |
2015 |
CIDR |
4.6176856e-05 |
| 7,956 |
Shasta: Interactive Reporting At Scale |
2016 |
SIGMOD |
4.6089395e-05 |
| 7,963 |
Building Community-Centric Information Exploration Applications on Social Content Sites |
2009 |
SIGMOD |
4.6089395e-05 |
| 8,398 |
Toward Progress Indicators on Steroids for Big Data Systems |
2013 |
CIDR |
4.5207516e-05 |
| 8,420 |
Handling Environments in a Nested Relational Algebra with Combinators and an Implementation in a Verified Query Compiler |
2017 |
SIGMOD |
4.5116638e-05 |
| 8,460 |
Piranha: Optimizing Short Jobs in Hadoop |
2013 |
VLDB |
4.5008938e-05 |
| 8,787 |
From SPARQL to MapReduce: The Journey Using a Nested TripleGroup Algebra |
2011 |
VLDB |
4.446583e-05 |
| 8,927 |
QMapper for Smart Grid: Migrating SQL-based Application to Hive |
2015 |
SIGMOD |
4.4229886e-05 |
| 8,985 |
SpongeFiles: Mitigating Data Skew in MapReduce Using Distributed Memory |
2014 |
SIGMOD |
4.4129993e-05 |
| 9,008 |
The Power of Nested Parallelism in Big Data Processing – Hitting Three Flies with One Slap – |
2021 |
SIGMOD |
4.4065351e-05 |
| 9,010 |
DataGarage: Warehousing Massive Performance Data on Commodity Servers |
2010 |
VLDB |
4.405974e-05 |
| 9,353 |
Rank Join Queries in NoSQL Databases |
2014 |
VLDB |
4.3485003e-05 |
| 9,384 |
Versatile Optimization of UDF-heavy Data Flows with Sofa |
2014 |
SIGMOD |
4.3432098e-05 |
| 9,520 |
PAXQuery: Parallel Analytical XML Processing |
2015 |
SIGMOD |
4.3282242e-05 |
| 9,613 |
Graft: A Debugging Tool For Apache Giraph |
2015 |
SIGMOD |
4.3136057e-05 |
| 11,199 |
QaaD (Query-as-a-Data): Scalable Execution of Massive Number of Small Queries in Spark |
2023 |
SIGMOD |
4.1905499e-05 |
| 11,215 |
Udon: Efficient Debugging of User-Defined Functions in Big Data Systems with Line-by-Line Control |
2023 |
SIGMOD |
4.1905499e-05 |
| 11,695 |
Integration of Large-Scale Data Processing Systems and Traditional Parallel Database Technology |
2019 |
VLDB |
4.1905499e-05 |
| 11,839 |
Logical Aspects of Massively Parallel and Distributed Systems |
2016 |
PODS |
4.1905499e-05 |
| 11,867 |
dmapply: A functional primitive to express distributed machine learning algorithms in R |
2016 |
VLDB |
4.1905499e-05 |
| 11,890 |
Parallel Evaluation of Multi-Semi-Joins |
2016 |
VLDB |
4.1905499e-05 |
| 11,898 |
Let's Rethink Join Optimization in Distributed Systems |
2015 |
CIDR |
4.1905499e-05 |
| 11,902 |
Building Highly-Optimized, Low-Latency Pipelines for Genomic Data Analysis |
2015 |
CIDR |
4.1905499e-05 |
| 11,924 |
A Demonstration of Rubato DB: A Highly Scalable NewSQL Database System for OLTP and Big Data Applications |
2015 |
SIGMOD |
4.1905499e-05 |
| 11,927 |
ShareInsights - An Unified Approach to Full-stack Data Processing |
2015 |
SIGMOD |
4.1905499e-05 |
| 11,984 |
Anti-Combining for MapReduce |
2014 |
SIGMOD |
4.1905499e-05 |
| 12,117 |
Declarative Error Management for Robust Data-Intensive Applications |
2012 |
SIGMOD |
4.1905499e-05 |