DBScholar

Back to authors

Samuel R. Madden

Author ID
o0000-0002-7470-3265
ORCID
0000-0002-7470-3265
Links
(found by gpt-5.6-luna on jul 24 2026)
Most Frequent Institution
Massachusetts Institute of Technology
Pagerank
1.4017771
Overall Rank
1 | 100.00%
Paper Count
146

Affiliation Timeline

Incoming Non-self Citations Over Time

Total yearly non-self incoming citations across all papers by this author.

Publications by Paper Pagerank

Showing all 146 publications. Total citations include self and non-self citations.

Rank Title Year Venue Total Citations Pagerank
12 C-Store: A Column-oriented DBMS 2005 VLDB 210 0.0006897844
43 A Comparison of Approaches to Large-Scale Data Analysis 2009 SIGMOD 71 0.0004552807
61 Integrating Compression and Execution in Column-Oriented Database Systems 2006 SIGMOD 128 0.00039236924
70 The End of an Architectural Era (It’s Time for a Complete Rewrite) 2007 VLDB 114 0.00037851432
80 H-Store: A High-Performance, Distributed Main Memory Transaction Processing System 2008 VLDB 126 0.00036354352
112 TelegraphCQ: Continuous Dataflow Processing for an Uncertain World 2003 CIDR 105 0.00032445088
123 Schism: a Workload-Driven Approach to Database Replication and Partitioning 2010 VLDB 111 0.00030749898
157 OLTP Through the Looking Glass, and What We Found There 2008 SIGMOD 72 0.00028316479
190 Scorpion: Explaining Away Outliers in Aggregate Queries 2013 VLDB 83 0.0002582857
200 Continuously Adaptive Continuous Queries over Streams 2002 SIGMOD 68 0.00025444144
257 Crowdsourced Databases: Query Processing with People 2011 CIDR 48 0.00022958404
266 Human-powered Sorts and Joins 2012 VLDB 42 0.00022735106
331 Column-Stores vs. Row-Stores: How Different Are They Really? 2008 SIGMOD 62 0.0002076806
343 Model-Driven Data Acquisition in Sensor Networks 2004 VLDB 59 0.00020510274
376 Scalable Semantic Web Data Management Using Vertical Partitioning 2007 VLDB 38 0.0001961765
395 SeeDB: Efficient Data-Driven Visualization Recommendations to Support Visual Analytics 2015 VLDB 53 0.00019156481
417 MauveDB: Supporting Model-based User Views in Database Systems 2006 SIGMOD 37 0.00018625112
420 Processing Analytical Queries over Encrypted Data 2013 VLDB 36 0.00018531799
431 HYRISE—A Main Memory Hybrid Storage Engine 2011 VLDB 54 0.00018400856
511 TelegraphCQ: Continuous Dataflow Processing 2003 SIGMOD 38 0.00017058115
555 SageDB: A Learned Database System 2019 CIDR 63 0.0001650754
585 Workload-Aware Database Monitoring and Consolidation 2011 SIGMOD 31 0.00015942562
634 Performance Tradeoffs in Read-Optimized Databases 2006 VLDB 31 0.00015367822
641 Relational Cloud: A Database-as-a-Service for the Cloud 2011 CIDR 30 0.00015248402
673 The TileDB Array Data Storage Manager 2017 VLDB 32 0.00014891415
717 Palimpzest: Optimizing AI-Powered Analytics with Declarative Query Processing 2025 CIDR 46 0.00014528323
797 Low Overhead Concurrency Control for Partitioned Main Memory Databases 2010 SIGMOD 38 0.00013916653
850 The Design of an Acquisitional Query Processor For Sensor Networks 2003 SIGMOD 33 0.00013482116
932 Starling: A Scalable Query Engine on Cloud Functions 2020 SIGMOD 38 0.00013002837
972 The Data Civilizer System 2017 CIDR 56 0.00012757732
1,023 MIRIS: Fast Object Track Queries in Video 2020 SIGMOD 37 0.00012430702
1,029 ZStream: A Cost-based Query Processor for Adaptively Detecting Composite Events 2009 SIGMOD 36 0.00012411098
1,086 DataHub: Collaborative Data Science & Dataset Version Management at Scale 2015 CIDR 31 0.0001209772
1,246 Blink and It's Done: Interactive Queries on Very Large Data 2012 VLDB 28 0.000113499
1,269 A Study of the Fundamental Performance Characteristics of GPUs and CPUs for Database Analytics 2020 SIGMOD 45 0.00011254742
1,387 Symphony: Towards Natural Language Query Answering over Multi-modal Data Lakes 2023 CIDR 19 0.00010829741
1,409 A Demonstration of SciDB: A Science-Oriented DBMS 2009 VLDB 17 0.00010738582
1,428 Knowing When You’re Wrong: Building Fast and Reliable Approximate Query Processing Systems 2014 SIGMOD 39 0.0001069161
1,457 Querying Continuous Functions in a Database System 2008 SIGMOD 16 0.00010585094
1,473 Voodoo - A Vector Algebra for Portable Database Performance on Modern Hardware 2016 VLDB 37 0.00010552993
1,493 Synthesizing Entity Matching Rules by Examples 2018 VLDB 31 0.00010500948
1,662 Rapid Sampling for Visualizations with Ordering Guarantees 2015 VLDB 31 9.9502569e-05
1,692 MISTIQUE: A System to Store and Query Model Intermediates for Model Diagnosis 2018 SIGMOD 28 9.8526476e-05
1,694 Aria: A Fast and Practical Deterministic OLTP Database 2020 VLDB 44 9.8520416e-05
1,805 Raha: A Configuration-Free Error Detection System 2019 SIGMOD 37 9.5938877e-05
1,829 Fault-Tolerance in the Borealis Distributed Stream Processing System 2005 SIGMOD 19 9.5470848e-05
1,833 MacroBase: Prioritizing Attention in Fast Data 2017 SIGMOD 31 9.5376219e-05
1,852 CrowdDB: Query Processing with the VLDB Crowd 2011 VLDB 14 9.4992704e-05
1,926 Counting with the Crowd 2013 VLDB 19 9.3636304e-05
1,939 S-Store: Streaming Meets Transaction Processing 2015 VLDB 22 9.3299482e-05
1,940 Performance and Resource Modeling in Highly-Concurrent OLTP Workloads 2013 SIGMOD 26 9.3298436e-05
1,993 RPT: Relational Pre-trained Transformer Is Almost All You Need towards Democratizing Data Preparation 2021 VLDB 27 9.2348951e-05
2,194 Demonstration of Qurk: A Query Processor for Human Operators 2011 SIGMOD 9 8.8780295e-05
2,241 Query Optimization for Dynamic Imputation 2017 VLDB 18 8.7704255e-05
2,251 Evaluating End-to-End Optimization for Data Analytics Applications in Weld 2018 VLDB 26 8.750953e-05
2,282 Decibel: The Relational Dataset Branching System 2016 VLDB 18 8.6997852e-05
2,326 Correlation Maps: A Compressed Access Method for Exploiting Soft Functional Dependencies 2009 VLDB 19 8.628257e-05
2,356 Vertexica: Your Relational Friend for Graph Analytics! 2014 VLDB 15 8.5852943e-05
2,363 Streaming Similarity Search over one Billion Tweets using Parallel Locality-Sensitive Hashing 2013 VLDB 12 8.572554e-05
2,416 Using Probabilistic Models for Data Management in Acquisitional Environments 2005 CIDR 12 8.4946354e-05
2,424 BigDansing: A System for Big Data Cleansing 2015 SIGMOD 34 8.483813e-05
2,467 Few-shot Text-to-SQL Translation using Structure and Content Prompt Learning 2023 SIGMOD 19 8.4171827e-05
2,492 A Demonstration of the BigDAWG Polystore System 2015 VLDB 13 8.3861255e-05
2,643 Scaling Up Crowd-Sourcing to Very Large Datasets: A Case for Active Learning 2015 VLDB 18 8.1741373e-05
2,709 An Integrated Approach to Recovery and High Availability in an Updatable, Distributed Data Warehouse 2006 VLDB 7 8.103163e-05
2,815 Distributing Queries over Low-Power Wireless Sensor Networks 2002 SIGMOD 4 7.9755556e-05
2,844 FactorJoin: A New Cardinality Estimation Framework for Join Queries 2023 SIGMOD 35 7.9446987e-05
2,957 Automatic Partitioning of Database Applications 2012 VLDB 16 7.808922e-05
2,981 StatusQuo: Making Familiar Abstractions Perform Using Program Analysis 2013 CIDR 10 7.7855712e-05
3,008 GenBase: A Complex Analytics Genomics Benchmark 2014 SIGMOD 11 7.757408e-05
3,049 Abacus: A Cost-Based Optimizer for Semantic Operator Systems 2026 VLDB 18 7.7087759e-05
3,059 The Case for RodentStore, an Adaptive, Declarative Storage System 2009 CIDR 13 7.6939894e-05
3,107 Top-k Queries on Uncertain Data: On Score Distribution and Typical Answers 2009 SIGMOD 9 7.640909e-05
3,160 S-Store: A Streaming NewSQL System for Big Velocity Applications 2014 VLDB 14 7.5773134e-05
3,327 Robust Query Driven Cardinality Estimation under Changing Workloads 2023 VLDB 37 7.4233639e-05
3,436 Collaborative Data Analytics with DataHub 2015 VLDB 8 7.297345e-05
3,489 Combining Small Language Models and Large Language Models for Zero-Shot NL2SQL 2024 VLDB 21 7.2594757e-05
3,500 CORADD: Correlation Aware Database Designer for Materialized Views and Indexes 2010 VLDB 10 7.2506844e-05
3,595 The Case for a Signal-Oriented Data Stream Management System 2007 CIDR 8 7.1790428e-05
3,657 Tile-based Lightweight Integer Compression in GPU 2022 SIGMOD 19 7.1227107e-05
3,684 Generating Concise Entity Matching Rules 2017 SIGMOD 9 7.0983735e-05
3,807 Vaas: Video Analytics At Scale 2020 VLDB 9 7.0081164e-05
3,925 Sloth: Being Lazy is a Virtue (When Issuing Database Queries) 2014 SIGMOD 10 6.919284e-05
3,947 OTIF: Efficient Tracker Pre-processing over Large Video Datasets 2022 SIGMOD 14 6.9052646e-05
3,956 Smile: A System to Support Machine Learning on EEG Data at Scale 2019 VLDB 6 6.8992907e-05
4,023 STAR: Scaling Transactions through Asymmetric Replication 2019 VLDB 20 6.8464253e-05
4,078 DBSeer: Resource and Performance Prediction for Building a Next Generation Database Cloud 2013 CIDR 17 6.8154962e-05
4,105 AutoOD: Automatic Outlier Detection 2023 SIGMOD 8 6.8033852e-05
4,534 AdaptDB: Adaptive Partitioning for Distributed Joins 2017 VLDB 13 6.5535468e-05
4,558 Databases Unbound: Querying All of the World’s Bytes with AI 2024 VLDB 7 6.5333125e-05
4,573 SILKMOTH: An Efficient Method for Finding Related Sets with Maximum Matching Constraints 2017 VLDB 11 6.5257221e-05
4,689 A Demonstration of AutoOD: A Self-Tuning Anomaly Detection System 2022 VLDB 6 6.4677473e-05
4,822 Epoch-based Commit and Replication in Distributed OLTP Databases 2021 VLDB 13 6.3931726e-05
4,915 REED: Robust, Efficient Filtering and Event Detection in Sensor Networks 2005 VLDB 5 6.3530622e-05
4,979 SeeDB: Automatically Generating Query Visualizations 2014 VLDB 9 6.3275644e-05
5,181 The Case for Data Visualization Management Systems 2014 VLDB 7 6.2377699e-05
5,255 Dagger: A Data (not code) Debugger 2020 CIDR 12 6.2057681e-05
5,501 Tweets as Data: Demonstration of TweeQL and TwitInfo 2011 SIGMOD 4 6.1018416e-05
5,703 A Demo of the Data Civilizer System 2017 SIGMOD 7 6.02892e-05
5,842 Pando: Enhanced Data Skipping with Logical Data Partitioning 2023 VLDB 7 5.971167e-05
5,929 Self-Organizing Data Containers 2022 CIDR 7 5.9411211e-05
5,935 Continuously Adaptive Similarity Search 2020 SIGMOD 5 5.9391038e-05
6,063 Speeding up Database Applications with Pyxis 2013 SIGMOD 3 5.8958547e-05
6,586 Extract-Transform-Load for Video Streams 2023 VLDB 9 5.741099e-05
6,627 A Demonstration of DBWipes: Clean as You Query 2012 VLDB 3 5.7281668e-05
6,635 What Makes a Good Physical Plan? — Experiencing Hardware-Conscious Query Optimization with Candomble 2016 SIGMOD 2 5.7243678e-05
6,715 Replicated Layout for In-Memory Database Systems 2022 VLDB 5 5.6989853e-05
6,989 Cackle: Analytical Workload Cost and Performance Stability With Elastic Pools 2023 SIGMOD 5 5.6257777e-05
7,038 CHIC: A Combination-based Recommendation System 2013 SIGMOD 1 5.6132544e-05
7,043 Human-in-the-loop Outlier Detection 2020 SIGMOD 17 5.611642e-05
7,068 Efficient Top-K Query Processing on Massively Parallel Hardware 2018 SIGMOD 5 5.6060783e-05
7,151 An Integration Framework for Sensor Networks and Data Stream Management Systems 2004 VLDB 3 5.5963104e-05
7,259 SeeSaw: Interactive Ad-hoc Search Over Image Databases 2023 SIGMOD 4 5.5695263e-05
7,293 Data Civilizer 2.0: A Holistic Framework for Data Preparation and Analytics 2019 VLDB 7 5.5616488e-05
7,503 Exploring big volume sensor data with Vroom 2017 VLDB 1 5.5057966e-05
7,708 UPI: A Primary Index for Uncertain Databases 2010 VLDB 3 5.4722487e-05
7,869 Blueprinting the Cloud: Unifying and Automatically Optimizing Cloud Data Infrastructures with BRAD 2024 VLDB 6 5.4342182e-05
7,879 Asynchronous Prefix Recoverability for Fast Distributed Stores 2021 SIGMOD 6 5.4332155e-05
7,928 Outlier Summarization via Human Interpretable Rules 2024 VLDB 4 5.4233854e-05
8,347 Efficient Discovery of Sequence Outlier Patterns 2019 VLDB 2 5.3487955e-05
8,451 CARTILAGE: Adding Flexibility to the Hadoop Skeleton 2013 SIGMOD 3 5.3324907e-05
8,549 Serverless State Management Systems 2024 CIDR 4 5.3164803e-05
8,643 PBench: Workload Synthesizer with Real Statistics for Cloud Analytics Benchmarking 2025 VLDB 2 5.2956539e-05
9,044 LANCET: Labeling Complex Data at Scale 2021 VLDB 3 5.2308178e-05
9,060 Optimizing Disjunctive Queries with Tagged Execution 2024 SIGMOD 2 5.227932e-05
9,124 Check Out the Big Brain on BRAD: Simplifying Cloud Data Processing with Learned Automated Data Meshes 2023 VLDB 6 5.2247088e-05
9,682 Debugging Large-Scale Data Science Pipelines using Dagger 2020 VLDB 3 5.1413762e-05
9,815 Lingua Manga: A Generic Large Language Model Centric System for Data Curation 2023 VLDB 1 5.1233734e-05
10,087 DARQ Matter Binds Everything: Performant and Composable Cloud Programming via Resilient Steps 2023 SIGMOD 3 5.0806786e-05
10,128 Amoeba: A Shape changing Storage System for Big Data 2016 VLDB 2 5.0729194e-05
10,152 Improving DBMS Scheduling Decisions with Accurate Performance Prediction on Concurrent Queries 2025 VLDB 1 5.0691578e-05
10,479 HotHash: Hotness-Aware Consistent Hashing for Cloud Databases 2026 SIGMOD 0 4.9769913e-05
10,989 Carnot: Interpretable, Interactive, and Optimized Execution of Deep Research Queries 2026 VLDB 0 4.9769913e-05
11,078 KEN: An Execution Engine for Unstructured Database Systems 2026 VLDB 0 4.9769913e-05
11,179 Virtualizing Cloud Data Infrastructures with BRAD 2025 SIGMOD 0 4.9769913e-05
11,519 RITA: Group Attention is All You Need for Timeseries Analytics 2024 SIGMOD 0 4.9769913e-05
11,562 MetaStore: Analyzing Deep Learning Meta-Data at Scale 2024 VLDB 0 4.9769913e-05
12,022 ATLANTIC: Making Database Differentially Private and Faster with Accuracy Guarantee 2021 VLDB 0 4.9769913e-05
12,193 SWIFT: Mining Representative Patterns from Large Event Streams 2019 VLDB 0 4.9769913e-05
12,654 No Bits Left Behind 2011 CIDR 0 4.9769913e-05
12,691 A Demonstration of HYRISE—A Main Memory Hybrid Storage Engine 2011 VLDB 0 4.9769913e-05
12,813 Demonstration of the TrajStore System 2009 VLDB 0 4.9769913e-05
12,839 How Best to Build Web-Scale Data Managers? A Panel Discussion 2009 VLDB 0 4.9769913e-05
13,979 MAQSA: A System for Social Analytics on News 2012 SIGMOD 0 -
14,085 Performance Profiling with EndoScope, an Acquisitional Software Monitoring Framework 2008 VLDB 0 -
14,132 Data Management in the CarTel Mobile Sensor Computing System 2006 SIGMOD 0 -

Frequent Co-authors

Co-authored at least 5 papers.

Co-author Shared Papers Rank Pagerank
Lei Cao 24 110 0.37918566
Mike Stonebraker 22 7 1.0496223
Tim Kraska 19 13 0.93045773
Nan Tang 13 41 0.5967
Daniel J. Abadi 13 58 0.51959547
Eugene Wu 12 61 0.50817259
Stanley Zdonik 10 52 0.53987344
Michael John Cafarella 10 80 0.44642508
Amol Deshpande 9 103 0.38441023
Mourad Ouzzani 8 123 0.35665042
Nesime Tatbul 8 151 0.31786478
Ziniu Wu 8 630 0.1067051
Yizhou Yan 8 1,233 0.061090718
Michael J. Franklin 7 14 0.92479545
Joseph M. Hellerstein 7 18 0.87118902
Elke Rundensteiner 7 72 0.47097977
Adam Marcus 7 917 0.079005256
Tianyu Li 7 1,127 0.065697422
Aditya Parameswaran 6 53 0.53819602
Carlo Curino 6 94 0.39977094
Alekh Jindal 6 182 0.2861186
David R. Karger 6 687 0.10051612
Hari Balakrishnan 6 863 0.083032679
Anil Shanbhag 6 885 0.08112168
Manasi Vartak 6 1,010 0.072018739
Guoliang Li 5 4 1.1617319
Ju Fan 5 134 0.33978843
Xiangyao Yu 5 167 0.3048865
Barzan Mozafari 5 208 0.25669786
Ahmed K. Elmagarmid 5 219 0.24491915
Alvin Cheung 5 233 0.23522504
Raul Castro Fernandez 5 264 0.21519947
Philippe Cudré-Mauroux 5 278 0.20426902
Alexander Rasin 5 386 0.16042669
Hideaki Kimura 5 663 0.10285595
Yi Lu 5 785 0.08965829
Evan P.C. Jones 5 860 0.083133659
Robert Miller 5 1,328 0.057140539
Zui Chen 5 1,758 0.045192763
Oscar Moll 5 1,981 0.041093212