SeGraM: a universal hardware accelerator for genomic sequence-to-graph and sequence-to-sequence mapping
Damla Senol Cali, Konstantinos Kanellopoulos, Joël Lindegger, Zülal Bingöl, Gurpreet S. Kalsi, Ziyi Zuo, Can Firtina, Meryem Banu Cavlak, Jeremie S. Kim, Nika Mansouri-Ghiasi, Gagandeep Singh, Juan Gómez-Luna
摘要
A critical step of genome sequence analysis is the mapping of sequenced DNA fragments (i.e., reads) collected from an individual to a known linear reference genome sequence (i.e., sequence-tosequence mapping). Recent works replace the linear reference sequence with a graph-based representation of the reference genome, which captures the genetic variations and diversity across many individuals in a population. Mapping reads to the graph-based reference genome (i.e., sequence-to-graph mapping) results in notable quality improvements in genome analysis. Unfortunately, while sequence-to-sequence mapping is well studied with many available tools and accelerators, sequence-to-graph mapping is a more difficult computational problem, with a much smaller number of practical software tools currently available.
We analyze two state-of-the-art sequence-to-graph mapping tools and reveal four key issues. We find that there is a pressing need to have a specialized, high-performance, scalable, and low-cost algorithm/hardware co-design that alleviates bottlenecks in both the seeding and alignment steps of sequence-to-graph mapping. Since sequence-to-sequence mapping can be treated as a special case of sequence-to-graph mapping, we aim to design an accelerator that is efficient for both linear and graph-based read mapping.
To this end, we propose SeGraM, a universal algorithm/hardware co-designed genomic mapping accelerator that can effectively and efficiently support both sequence-to-graph mapping and sequenceto-sequence mapping, for both short and long reads. To our knowledge, SeGraM is the first algorithm/hardware co-design for accelerating sequence-to-graph mapping. SeGraM consists of two main components: (1) MinSeed, the first minimizer-based seeding accelerator, which finds the candidate locations in a given genome graph; and (2) BitAlign, the first bitvector-based sequence-to-graph alignment accelerator, which performs alignment between a given read and the subgraph identified by MinSeed. We couple SeGraM with high-bandwidth memory to exploit low latency and highlyparallel memory access, which alleviates the memory bottleneck.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- GenDP: A Framework of Dynamic Programming Acceleration for Genome Sequencing AnalysisYufeng Gu, Arun Subramaniyan, Timothy Dunn, Alireza Khadem 等ISCA 2023 · 被引用 19 次
- TALCO: Tiling Genome Sequence Alignment Using Convergence of Traceback PointersSumit Walia, Cheng Ye, Arkid Bera, Dhruvi Lodhavia 等HPCA 2024 · 被引用 14 次
- SMX: Heterogeneous Architecture for Universal Sequence Alignment AccelerationMax Doblas, Po Jui Shih, Oscar Lostes-Cazorla, Miquel Moretó 等MICRO 2025 · 被引用 7 次
- Rapid GPU-Based Pangenome Graph LayoutJiajie Li, Jan-Niklas Schmelzle, Yixiao Du, Simon Heumos 等SC 2024 · 被引用 4 次
- SAGe: A Lightweight Algorithm-Architecture Co-Design for Mitigating the Data Preparation Bottleneck in Large-Scale Genome Sequence AnalysisNika Mansouri-Ghiasi, Talu Güloglu, Harun Mustafa, Can Firtina 等HPCA 2026 · 被引用 3 次
它引用的顶会 Paper9
- SISA: Set-Centric Instruction Set Architecture for Graph Mining on Processing-in-Memory SystemsMaciej Besta, Raghavendra Kanakagiri, Grzegorz Kwasniewski, Rachata Ausavarungnirun 等MICRO 2021 · 被引用 78 次
- GenStore: a high-performance in-storage processing system for genome sequence analysisNika Mansouri-Ghiasi, Jisung Park, Harun Mustafa, Jeremie S. Kim 等ASPLOS 2022 · 被引用 74 次
- SquiggleFilter: An Accelerator for Portable Virus DetectionTimothy Dunn, Harisankar Sadasivan, Jack Wadden, Kush Goliya 等MICRO 2021 · 被引用 61 次
- SeedEx: A Genome Sequencing Accelerator for Optimal Alignments in Subminimal SpaceDaichi Fujiki, Shunhao Wu, Nathan Ozog, Kush Goliya 等MICRO 2020 · 被引用 52 次
- Sieve: Scalable In-situ DRAM-based Accelerator Designs for Massively Parallel k-mer MatchingLingxi Wu, Rasool Sharifi, Marzieh Lenjani, Kevin Skadron 等ISCA 2021 · 被引用 31 次
相关 Paper
- Harp: Leveraging Quasi-Sequential Characteristics to Accelerate Sequence-to-Graph Mapping of Long ReadsYichi Zhang, Dibei Chen, Gang Zeng, Jianfeng Zhu 等ASPLOS 2024 · 被引用 5 次
- GenASM: A High-Performance, Low-Power Approximate String Matching Acceleration Framework for Genome Sequence AnalysisDamla Senol Cali, Gurpreet S. Kalsi, Zülal Bingöl, Can Firtina 等MICRO 2020 · 被引用 23 次
- GenPairX: A Hardware-Algorithm Co-Designed Accelerator for Paired-End Read MappingJulien Eudine, Chu Li, Zhuo Cheng, Renzo Andri 等HPCA 2026 · 被引用 2 次
- MeG2: In-Memory Acceleration for Genome Graphs AnalysisYu Huang, Long Zheng, Haifeng Liu, Zhuoran Zhou 等DAC 2023 · 被引用 3 次
- NvWa: Enhancing Sequence Alignment Accelerator Throughput via Hardware SchedulingYewen Li, Xueqi Li, Ruihao Gao, Wanqi Liu 等HPCA 2023 · 被引用 6 次
