A compressive seeding algorithm in conjunction with reordering-based compression
文献类型: 外文期刊
第一作者: Ji, Fahu
作者: Ji, Fahu;Liu, Xianming;Zhou, Qian;Liu, Xianming;Ruan, Jue;Zhu, Zexuan
作者机构:
期刊名称:BIOINFORMATICS ( 影响因子:5.8; 五年影响因子:8.3 )
ISSN: 1367-4803
年卷期: 2024 年 40 卷 3 期
页码:
收录情况: SCI
摘要: Motivation Seeding is a rate-limiting stage in sequence alignment for next-generation sequencing reads. The existing optimization algorithms typically utilize hardware and machine-learning techniques to accelerate seeding. However, an efficient solution provided by professional next-generation sequencing compressors has been largely overlooked by far. In addition to achieving remarkable compression ratios by reordering reads, these compressors provide valuable insights for downstream alignment that reveal the repetitive computations accounting for more than 50% of seeding procedure in commonly used short read aligner BWA-MEM at typical sequencing coverage. Nevertheless, the exploited redundancy information is not fully realized or utilized.Results In this study, we present a compressive seeding algorithm, named CompSeed, to fill the gap. CompSeed, in collaboration with the existing reordering-based compression tools, finishes the BWA-MEM seeding process in about half the time by caching all intermediate seeding results in compact trie structures to directly answer repetitive inquiries that frequently cause random memory accesses. Furthermore, CompSeed demonstrates better performance as sequencing coverage increases, as it focuses solely on the small informative portion of sequencing reads after compression. The innovative strategy highlights the promising potential of integrating sequence compression and alignment to tackle the ever-growing volume of sequencing data.Availability and implementation CompSeed is available at https://github.com/i-xiaohu/CompSeed.
分类号:
- 相关文献
 
作者其他论文 更多>>
- 
                                    
Decoding the fish genome opens a new era in important trait research and molecular breeding in China
作者:Zhou, Qian;Wang, Jialin;Chen, Zhangfan;Wang, Na;Li, Ming;Wang, Lei;Si, Yufeng;Lu, Sheng;Cui, Zhongkai;Liu, Xuhui;Chen, Songlin;Zhou, Qian;Wang, Jialin;Chen, Zhangfan;Wang, Na;Li, Ming;Wang, Lei;Si, Yufeng;Lu, Sheng;Cui, Zhongkai;Liu, Xuhui;Chen, Songlin;Li, Jiongtang
关键词:fish genomics; economic traits; genomic selection; genome editing; genomic breeding
 - 
                                    
Garlic peel extract as an antioxidant inhibits triple-negative breast tumor growth and angiogenesis by inhibiting cyclooxygenase-2 expression
作者:Dong, Yushi;Yue, Xiqing;Li, Mohan;Zhou, Qian;Zhang, Jiyue;Xie, Aijun
关键词:antioxidant; COX-2; garlic peel; triple-negative breast cancer
 - 
                                    
Low-input PacBio sequencing generates high-quality individual fly genomes and characterizes mutational processes
作者:Jia, Hangxing;Tan, Shengjun;Cai, Yingao;Guo, Yanyan;Shen, Jieyu;Zhang, Yaqiong;Ma, Huijing;Zhang, Qingzhu;Qiao, Gexia;Zhang, Yong E.;Cai, Yingao;Guo, Yanyan;Shen, Jieyu;Zhang, Qingzhu;Chen, Jinfeng;Qiao, Gexia;Zhang, Yong E.;Chen, Jinfeng;Ruan, Jue
关键词:
 - 
                                    
Rhodopseudomonas palustris shapes bacterial community, reduces Cd bioavailability in Cd contaminated flooding paddy soil, and improves rice performance
作者:Su, Yanqiu;Su, Yanqiu;Li, Ziyuan;Deng, Hongmei;Zhou, Qian;Li, Lihuan;Zhao, Lanyin;Shi, Qiuyun;Chen, Yanger;Yuan, Shu;Liu, Qi
关键词:Photosynthetic bacteria; Rhodopseudomonas palustris; Cadmium; Oryza sativa L.; Anaerobic bacteria
 - 
                                    
HiTE: a fast and accurate dynamic boundary adjustment approach for full-length transposable element detection and annotation
作者:Hu, Kang;Ni, Peng;Xu, Minghua;Zou, You;Wang, Jianxin;Hu, Kang;Ni, Peng;Wang, Jianxin;Hu, Kang;Ni, Peng;Xu, Minghua;Zou, You;Wang, Jianxin;Chang, Jianye;Ruan, Jue;Gao, Xin;Gao, Xin;Li, Yaohang;Hu, Bin;Hu, Bin
关键词:
 - 
                                    
A broad host phage, CP6, for combating multidrug-resistant Campylobacter prevalent in poultry meat
作者:Zhang, Xiaoyan;Tang, Mengjun;Zhou, Qian;Lu, Junxian;Tang, Xiujun;Ma, Lina;Zhang, Jing;Chen, Dawei;Gao, Yushi;Zhang, Hui
关键词:Campylobacter; lytic phage; poultry meat; food safety
 - 
                                    
NextDenovo: an efficient error correction and accurate assembly tool for noisy long reads
作者:Hu, Jiang;Wang, Zhuo;Sun, Zongyi;Liang, Fan;Li, Jingjing;Wang, Depeng;Hu, Benxia;Ayoola, Adeola Oluwakemi;Wu, Dong-Dong;Wang, Sheng;Sandoval, Jose R.;Cooper, David N.;Hu, Jiang;Ye, Kai;Ruan, Jue;Xiao, Chuan-Le;Wu, Dong-Dong;Wu, Dong-Dong;Wang, Sheng;Wu, Dong-Dong
关键词:Long reads; Genome assembly; Error-correction; Human genomes; Segmental duplication