Title :
SRIdent: A novel pipeline for real-time identification of species from high-throughput sequencing reads in Metagenomics and clinical diagnostic assays
Author :
Ramin Karimi;Andras Hajdu
Author_Institution :
Faculty of Informatics, Department of Computer Graphics and Image Processing, University of Debrecen, 4032, Egyetem ter 1, Hungary
Abstract :
New advances in rapid sequencing of large amounts of DNA have brought a great potential for the study of complex communities of microorganisms. One of the challenging problems is rapid identification of species from sequenced reads. Delays in the identification of pathogens are a barrier to the early diagnosis and proper treatment of infectious diseases. In this paper we proposed SRIdent (Short Read Identifier), an effective pipeline for real-time identification of species from high-throughput sequencing reads in Metagenomics and clinical diagnostic assays. This pipeline is based on generating k-mers from the short reads and searching the existence of DNA signatures in the Reads k-mers, by using Apache Hive data-warehousing. RkmerG (Read k-mers Generator) is a software program presented in this paper, for producing k-mers of the short reads, in order to use in the pipeline. The purpose of this study is to identify the species in a sample, directly from the reads without assembling and alignment.
Keywords :
"DNA","Sequential analysis","Pipelines","Microorganisms","Genomics","Bioinformatics","Databases"
Conference_Titel :
Engineering in Medicine and Biology Society (EMBC), 2015 37th Annual International Conference of the IEEE
Electronic_ISBN :
1558-4615
DOI :
10.1109/EMBC.2015.7319877