DBScholar

Back to authors

Michael J. Franklin

Author ID
u62
ORCID
-
Links
(found by gpt-5.6-luna on jul 24 2026)
Most Frequent Institution
University of California Berkeley
Pagerank
0.92479545
Overall Rank
14 | 99.94%
Paper Count
104

Affiliation Timeline

Incoming Non-self Citations Over Time

Total yearly non-self incoming citations across all papers by this author.

Publications by Paper Pagerank

Showing all 104 publications. Total citations include self and non-self citations.

Rank Title Year Venue Total Citations Pagerank
23 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 219 0.00055384955
92 CrowdDB: Answering Queries with Crowdsourcing 2011 SIGMOD 77 0.00034670735
112 TelegraphCQ: Continuous Dataflow Processing for an Uncertain World 2003 CIDR 105 0.00032445088
198 CrowdER: Crowdsourcing Entity Resolution 2012 VLDB 71 0.00025546182
247 Efficient Filtering of XML Documents for Selective Dissemination of Information 2000 VLDB 40 0.00023188523
302 Shoring Up Persistent Applications 1994 SIGMOD 57 0.00021664369
410 Semantic Data Caching and Replacement 1996 VLDB 42 0.00018796567
423 Cost-based Query Scrambling for Initial Delays 1998 SIGMOD 31 0.00018487497
432 Shark: SQL and Rich Analytics at Scale 2013 SIGMOD 54 0.00018331051
441 A Fast Index for Semistructured Data 2001 VLDB 27 0.00018206766
483 ActiveClean: Interactive Data Cleaning For Statistical Modeling 2016 VLDB 53 0.00017584249
511 TelegraphCQ: Continuous Dataflow Processing 2003 SIGMOD 38 0.00017058115
537 MLbase: A Distributed Machine-learning System 2013 CIDR 34 0.00016762761
717 Palimpzest: Optimizing AI-Powered Analytics with Declarative Query Processing 2025 CIDR 46 0.00014528323
808 Coordination Avoidance in Database Systems 2015 VLDB 37 0.00013772337
850 The Design of an Acquisitional Query Processor For Sensor Networks 2003 SIGMOD 33 0.00013482116
852 Streaming Queries over Streaming Data 2002 VLDB 29 0.00013450014
871 Leveraging Transitive Relations for Crowdsourced Joins 2013 SIGMOD 35 0.00013338722
888 Pay-as-you-go User Feedback for Dataspace Systems 2008 SIGMOD 29 0.00013250185
903 Towards an Internet-Scale XML Dissemination Service 2004 VLDB 19 0.00013171589
1,036 Fine-grained Partitioning for Aggressive Data Skipping 2014 SIGMOD 41 0.00012372946
1,114 Principles of Dataspace Systems 2006 PODS 28 0.00011959287
1,358 Design Considerations for High Fan-in Systems: The HiFi Approach 2005 CIDR 17 0.00010917457
1,361 Adaptive Cleaning for RFID Data Streams 2006 VLDB 20 0.00010912941
1,383 On-the-Fly Sharing for Streamed Aggregation 2006 SIGMOD 27 0.00010842596
1,629 SAND: Streaming Subsequence Anomaly Detection 2021 VLDB 37 0.0001003165
1,722 A Sample-and-Clean Framework for Fast and Accurate Query Processing on Dirty Data 2014 SIGMOD 38 9.7921604e-05
1,796 Dynamic Pipeline Scheduling for Improving Interactive Query Performance 2001 VLDB 14 9.6147923e-05
1,817 Broadcast Disks: Data Management for Asymmetric Communication Environments 1995 SIGMOD 14 9.5704315e-05
1,847 Data Market Platforms: Trading Data Assets to Solve Data Problems 2020 VLDB 22 9.5091361e-05
1,852 CrowdDB: Query Processing with the VLDB Crowd 2011 VLDB 14 9.4992704e-05
1,938 TSB-UAD: An End-to-End Benchmark Suite for Univariate Time-Series Anomaly Detection 2022 VLDB 38 9.334286e-05
1,987 Continuous Analytics Over Discontinuous Streams 2010 SIGMOD 15 9.2478037e-05
2,009 PIQL: Success-Tolerant Query Processing in the Cloud 2012 VLDB 19 9.1949132e-05
2,233 Shark: Fast Data Analysis Using Coarse-grained Distributed Memory 2012 SIGMOD 12 8.7921042e-05
2,383 Query Processing for High-Volume XML Message Brokering 2003 VLDB 13 8.5421855e-05
2,386 Performance Tradeoffs for Client-Server Query Processing 1996 SIGMOD 17 8.5358065e-05
2,431 Data Caching Tradeoffs in Client-Server DBMS Architectures 1991 SIGMOD 14 8.4760078e-05
2,643 Scaling Up Crowd-Sourcing to Very Large Datasets: A Case for Active Learning 2015 VLDB 18 8.1741373e-05
2,708 Probabilistically Bounded Staleness for Practical Partial Quorums 2012 VLDB 14 8.1033264e-05
2,782 Scheduling for shared window joins over data streams 2003 VLDB 17 8.0214634e-05
2,986 A Database Striptease or How to Manage Your Personal Databases 2003 VLDB 4 7.7772018e-05
2,992 Balancing Push and Pull for Data Broadcast 1997 SIGMOD 8 7.7700264e-05
3,083 Skipping-oriented Partitioning for Columnar Layouts 2017 VLDB 20 7.6620866e-05
3,200 Feral Concurrency Control: An Empirical Investigation of Modern Application Integrity 2015 SIGMOD 13 7.5406953e-05
3,236 How Large Language Models Will Disrupt Data Management 2023 VLDB 13 7.4996147e-05
3,299 Volume Under the Surface: A New Accuracy Evaluation Measure for Time-Series Anomaly Detection 2022 VLDB 25 7.441222e-05
3,391 RTP: Robust Tenant Placement for Elastic In-Memory Database Clusters 2013 SIGMOD 8 7.3462108e-05
3,516 GridDB: A Data-Centric Overlay for Scientific Grids 2004 VLDB 4 7.2392707e-05
3,532 Rethinking Data-Intensive Science Using Scalable Analytics Systems 2015 SIGMOD 4 7.2228355e-05
3,919 Fine-Grained Sharing in a Page Server OODBMS 1994 SIGMOD 6 6.9230012e-05
4,025 CLAMShell: Speeding up Crowds for Low-latency Data Labeling 2016 VLDB 12 6.8439934e-05
4,057 The Case for Precision Sharing 2004 VLDB 9 6.8230987e-05
4,143 Hybrid In-Database Inference for Declarative Information Extraction 2011 SIGMOD 8 6.7783148e-05
4,236 Efficient Incremental Garbage Collection for Client-Server Object Database Systems 1995 VLDB 4 6.7093088e-05
4,300 The Missing Piece in Complex Analytics: Low Latency, Scalable Model Management and Serving with Velox 2015 CIDR 11 6.676283e-05
4,317 GRAIL: Efficient Time-Series Representation Learning 2019 VLDB 26 6.6651881e-05
4,322 Generalized Scale Independence Through Incremental Precomputation 2013 SIGMOD 7 6.6629734e-05
4,355 Interaction of Query Evaluation and Buffer Management for Information Retrieval 1998 SIGMOD 2 6.6381897e-05
4,558 Databases Unbound: Querying All of the World’s Bytes with AI 2024 VLDB 7 6.5333125e-05
4,688 Continuous Analytics: Rethinking Query Processing in a Network-Effect World 2009 CIDR 6 6.4683444e-05
4,705 Debunking Four Long-Standing Misconceptions of Time-Series Distance Measures 2020 SIGMOD 24 6.4591377e-05
4,879 Events on the Edge 2005 SIGMOD 3 6.3692634e-05
4,962 PrivateClean: Data Cleaning and Differential Privacy 2016 SIGMOD 4 6.3360478e-05
4,976 Data Station: Delegated, Trustworthy, and Auditable Computation to Enable Data-Sharing Consortia with a Data Escrow 2022 VLDB 5 6.3288989e-05
5,024 Querying Probabilistic Information Extraction 2010 VLDB 7 6.3081571e-05
5,204 Global Memory Management in Client-Server DBMS Architectures 1992 VLDB 5 6.2273985e-05
5,207 Crash Recovery in Client-Server EXODUS 1992 SIGMOD 14 6.2260329e-05
5,470 Crowdsourcing Applications and Platforms: A Data Management Perspective 2011 VLDB 4 6.1173893e-05
5,526 A First Tutorial on Dataspaces 2008 VLDB 4 6.0922984e-05
5,615 CrowdQ: Crowdsourced Query Understanding 2013 CIDR 2 6.0640551e-05
5,871 PIQL: A Performance Insightful Query Language 2010 SIGMOD 3 5.9620208e-05
5,882 ActiveClean: An Interactive Data Cleaning Framework For Modern Machine Learning 2016 SIGMOD 6 5.9560519e-05
6,045 Remembrance of Streams Past: Overload-Sensitive Management of Archived Streams 2004 VLDB 4 5.9033006e-05
6,192 VergeDB: A Database for IoT Analytics on Edge Devices 2021 CIDR 26 5.8540507e-05
6,231 Crowds, Clouds, and Algorithms: Exploring the Human Side of "Big Data" Applications 2010 SIGMOD 3 5.8420493e-05
6,514 PLANET: Making Progress with Commit Processing in Unpredictable Environments 2014 SIGMOD 3 5.7570121e-05
6,801 Intermittent Query Processing 2019 VLDB 11 5.6776677e-05
6,820 CrocodileDB: Efficient Database Execution through Intelligent Deferment 2020 CIDR 6 5.6708756e-05
6,978 SparkR: Scaling R Programs with Spark 2016 SIGMOD 6 5.6272603e-05
7,543 HiFi: A Unified Architecture for High Fan-in Systems (System Demonstration) 2004 VLDB 4 5.4970407e-05
7,667 CYADB: A Database that Covers Your Ask 2018 VLDB 1 5.4746904e-05
7,734 Disseminating Updates on Broadcast Disks 1996 VLDB 3 5.4633791e-05
8,137 Thrifty Query Execution via Incrementability 2020 SIGMOD 6 5.391109e-05
8,172 Fast and Reliable Missing Data Contingency Analysis with Predicate-Constraints 2020 SIGMOD 5 5.3824434e-05
8,322 Adaptive Execution of Variable-Accuracy Functions 2006 VLDB 1 5.3534332e-05
8,785 Wisteria: Nurturing Scalable Data Cleaning Infrastructure 2015 VLDB 8 5.2767488e-05
8,876 Stale View Cleaning: Getting Fresh Answers from Stale Materialized Views 2015 VLDB 10 5.2587627e-05
9,329 CrocodileDB in Action: Resource-Efficient Query Execution by Exploiting Time Slackness 2020 VLDB 1 5.1921674e-05
9,633 Theseus: Navigating the Labyrinth of Time-Series Anomaly Detection 2022 VLDB 14 5.1459383e-05
9,844 "Data In Your Face": Push Technology in Perspective 1998 SIGMOD 3 5.1213404e-05
10,228 Towards Resource-adaptive Query Execution in Cloud Native Databases 2024 CIDR 2 5.054828e-05
11,755 Accelerating Similarity Search for Elastic Measures: A Study and New Generalization of Lower Bounding Distances 2023 VLDB 11 4.9769913e-05
12,488 A Partitioning Framework for Aggressive Data Skipping 2014 VLDB 0 4.9769913e-05
13,024 Predicate Result Range Caching for Continuous Queries 2005 SIGMOD 2 4.9769913e-05
13,110 GridDB: A Relational Interface for the Grid 2003 SIGMOD 1 4.9769913e-05
13,370 Local Disk Caching for Client-Server Database Systems 1993 VLDB 1 4.9769913e-05
13,719 Will LLMs reshape, supercharge, or kill data science? (VLDB 2023 Panel) 2023 VLDB 0 -
13,782 SAND in Action: Subsequence Anomaly Detection for Streams 2021 VLDB 16 -
13,921 Should we all be teaching “Intro to Data Science” instead of “Intro to Databases”? 2014 SIGMOD 0 -
13,946 PBS at Work: Advancing Data Management with Consistency Metrics 2013 SIGMOD 0 -
14,212 Rethinking the Conference Reviewing Process 2004 SIGMOD 1 -
14,360 Data Staging for On-Demand Broadcast 2001 VLDB 1 -
14,451 DBIS-Toolkit: Adaptable Middleware For Large Scale Data Delivery 1999 SIGMOD 2 -

Frequent Co-authors

Co-authored at least 5 papers.

Co-author Shared Papers Rank Pagerank
Tim Kraska 17 13 0.93045773
Sanjay Krishnan 17 250 0.22545869
Aaron J. Elmore 14 138 0.33483585
Joseph M. Hellerstein 10 18 0.87118902
Jiannan Wang 10 143 0.33012272
John Paparrizos 9 282 0.20276896
Sailesh Krishnamurthy 9 373 0.16622831
Reynold Xin 8 217 0.24640303
Samuel R. Madden 7 1 1.4017771
Eugene Wu 7 61 0.50817259
Ion Stoica 7 139 0.33329441
Mike Carey 6 21 0.77403259
Stanley Zdonik 6 52 0.53987344
Zechao Shang 6 640 0.10507889
Alon Y. Levy 5 25 0.7417918
Themistoklis Palpanas 5 82 0.43971713
Peter D. Bailis 5 133 0.34085096
Ali Ghodsi 5 425 0.14840305
Paul Boniol 5 788 0.089249623
Wei Hong 5 826 0.086499485
Shawn R. Jeffery 5 964 0.07567992
Ken Goldberg 5 1,705 0.046085305