The SMEM Seeding Acceleration for DNA Sequence Alignment

Mau-Chung Chang, Yu Ting Chen, Jason Cong, Po-Tsang Huang, Chun Liang Kuo, Cody Hao Yu

研究成果: Conference contribution同行評審

19 引文 斯高帕斯(Scopus)

摘要

The advance of next-generation sequencing technology has dramatically reduced the cost of genome sequencing. However, processing and analyzing huge amounts of data collected from sequencers introduces significant computation challenges, these have become the bottleneck in many research and clinical applications. For such applications, read alignment is usually one of the most compute-intensive steps. Billions of reads generated from the sequencer need to be aligned to the long reference genome. Recent state-of-the-art software read aligners follow the seed-andextend model. In this paper we focus on accelerating the first seeding stage, which generates the seeds using the supermaximal exact match (SMEM) seeding algorithm. The two main challenges for accelerating this process are 1) how to process a huge number of short reads with high throughput, and 2) how to hide the frequent and long random memory access when we try to fetch the value of the reference genome. In this paper, we propose a scalable array-based architecture, which is composed by many processing engines (PEs) to process large amounts of data simultaneously for the demand of high throughput. Furthermore, we provide a tight software/hardware integration that realizes the proposed architecture on the Intel-Altera HARP system. With a 16-PE accelerator engine, we accelerate the SMEM algorithm by 4x, and the overall SMEM seeding stage by 26% when compared with 16-thread CPU execution. We further analyze the performance bottleneck of the design due to extensive DRAM accesses and discuss the possible improvements that are worthwhile to be explored in the future.

原文English
主出版物標題Proceedings - 24th IEEE International Symposium on Field-Programmable Custom Computing Machines, FCCM 2016
發行者Institute of Electrical and Electronics Engineers Inc.
頁面32-39
頁數8
ISBN(電子)9781509023561
DOIs
出版狀態Published - 16 八月 2016
事件24th IEEE International Symposium on Field-Programmable Custom Computing Machines, FCCM 2016 - Washington, United States
持續時間: 1 五月 20163 五月 2016

出版系列

名字Proceedings - 24th IEEE International Symposium on Field-Programmable Custom Computing Machines, FCCM 2016

Conference

Conference24th IEEE International Symposium on Field-Programmable Custom Computing Machines, FCCM 2016
國家United States
城市Washington
期間1/05/163/05/16

指紋 深入研究「The SMEM Seeding Acceleration for DNA Sequence Alignment」主題。共同形成了獨特的指紋。

引用此