WHAM: A High-throughput Sequence Alignment Method
Summary: WHAM: high-throughput short-read alignment using hash-based indexing and bitwise ops, enabling richer match models than prior aligners. Significant speedups over state-of-the-art; code at http://www.cs.wisc.edu/wham/. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Yinan Li (University of Wisconsin)
- 2. Allison Terrell (University of Wisconsin)
- 3. Jignesh M. Patel (University of Wisconsin)
BibTeX Citation
@inproceedings{li_sigmod11,
title = {{WHAM: A High-throughput Sequence Alignment Method}},
author = {Li, Yinan and Terrell, Allison and Patel, Jignesh M.},
series = {{SIGMOD} '11},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/1989323.1989370},
url = {https://dl.acm.org/doi/10.1145/1989323.1989370},
year = {2011}
}
Incoming Citations (Sorted by Pagerank)
Showing 8 of 8 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 2,362 | Streaming Similarity Search over one Billion Tweets using Parallel Locality-Sensitive Hashing | 2013 | VLDB | 8.5746436e-05 |
| 3,532 | Rethinking Data-Intensive Science Using Scalable Analytics Systems | 2015 | SIGMOD | 7.2261031e-05 |
| 7,006 | A Generic Framework for Efficient and Effective Subsequence Retrieval | 2012 | VLDB | 5.6234186e-05 |
| 8,587 | Building Highly-Optimized, Low-Latency Pipelines for Genomic Data Analysis | 2015 | CIDR | 5.3078077e-05 |
| 10,298 | Local Filtering: Improving the Performance of Approximate Queries on String Collections | 2015 | SIGMOD | 5.0418674e-05 |
| 12,292 | Massively Parallel Processing of Whole Genome Sequence Data: An In-Depth Performance Study | 2017 | SIGMOD | 4.9793485e-05 |
| 12,574 | RCSI: Scalable similarity search in thousand(s) of genomes | 2013 | VLDB | 4.9793485e-05 |
| 12,626 | Massive Genomic Data Processing and Deep Analysis | 2012 | VLDB | 4.9793485e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 7 of 7 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 168 | Efficient Exact Set-Similarity Joins | 2006 | VLDB | 0.00027163517 |
| 207 | Cache Conscious Indexing for Decision-Support in Main Memory | 1999 | VLDB | 0.00024970987 |
| 1,061 | VGRAM: Improving Performance of Approximate Queries on String Collections Using Variable-Length Grams | 2007 | VLDB | 0.00012213729 |
| 2,144 | Cost-Based Variable-Length-Gram Selection for String Collections to Support Approximate Queries Efficiently | 2008 | SIGMOD | 8.9608583e-05 |
| 2,313 | n-Gram/2L: A Space and Time Efficient Two-Level n-Gram Inverted Index Structure | 2005 | VLDB | 8.6550779e-05 |
| 5,818 | Fast nGram-Based String Search Over Data Encoded Using Algebraic Signatures | 2007 | VLDB | 5.9837891e-05 |
| 6,305 | Reference-Based Alignment in Large Sequence Databases | 2009 | VLDB | 5.818151e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 13,969 | Memory Efficient Minimum Substring Partitioning | 2013 | VLDB |
| 2 | 12,292 | Massively Parallel Processing of Whole Genome Sequence Data: An In-Depth Performance Study | 2017 | SIGMOD |
| 3 | 6,619 | Reference-Based Indexing of Sequence Databases | 2006 | VLDB |
| 4 | 12,775 | Data Management for High-Throughput Genomics | 2009 | CIDR |
| 5 | 10,478 | LSHAlign: All-Pair Near-Duplicate Text Alignment via LSH | 2026 | SIGMOD |
| 6 | 5,600 | Approximate Encoding for Direct Access and Query Processing over Compressed Bitmaps | 2006 | VLDB |
| 7 | 6,305 | Reference-Based Alignment in Large Sequence Databases | 2009 | VLDB |
| 8 | 8,961 | ALAE: Accelerating Local Alignment with Affine Gap Exactly in Biosequence Databases | 2012 | VLDB |
| 9 | 12,626 | Massive Genomic Data Processing and Deep Analysis | 2012 | VLDB |
| 10 | 4,734 | Fast Processing and Querying of 170TB of Genomics Data via a Repeated And Merged BloOm Filter (RAMBO) | 2021 | SIGMOD |