Scalable Cloud-Based Data Analysis Software Systems for Big Data from Next Generation Sequencing

Monika Szczerba , Marek Wiewiórka , Michał Okoniewski , Henryk Rybiński


Next generation sequencing (NGS) technology has become a serious computational challenge since its commercial introduction in 2008. Currently, thousands of machines worldwide produce daily billions of sequenced nucleotide base pairs of data. Due to continuous development of faster and economical sequencing technologies, processing the large amounts of data produced by high throughput sequencing technologies became the main challenge in bioinformatics. It can be solved by the new generation of software tools based on the paradigms and principles developed within the Hadoop ecosystem. This chapter presents the overall perspective for data analysis software for genomics and prospects for the emerging applications. To show genomic big data analysis in practice, a case study of the SparkSeq system that delivers tool for biological sequence analysis is presented.
Author Monika Szczerba (FEIT / IN)
Monika Szczerba,,
- The Institute of Computer Science
, Marek Wiewiórka (FEIT / IN)
Marek Wiewiórka,,
- The Institute of Computer Science
, Michał Okoniewski (FEIT / IN) - [Research Informatics, Scientific IT Services, ETH Zürich, Zurich, Switzerland]
Michał Okoniewski,,
- The Institute of Computer Science
- Research Informatics, Scientific IT Services, ETH Zürich, Zurich, Switzerland
, Henryk Rybiński (FEIT / IN)
Henryk Rybiński,,
- The Institute of Computer Science
Publication size in sheets1
Book Japkowicz Nathalie, Stefanowski Jerzy (eds.): Big Data Analysis: New Algorithms for a New Society, Studies in Big Data, vol. 16, 2016, Heidelberg New York Dordrecht London, Springer International Publishing, ISBN 978-3-319-26987-0, [978-3-319-26989-4], 329 p., DOI:10.1007/978-3-319-26989-4
front_matter_BDANANS.pdf / 87.16 KB / No licence information
Keywords in EnglishGenomics – Big data – RNA – DNA – Next-generation sequencing – Biobanking
ProjectDevelopment of new algorithms in the areas of software and computer architecture, artificial intelligence and information systems and computer graphics . Project leader: Rybiński Henryk, , Phone: +48 22 234 7731, start date 18-05-2015, end date 30-11-2016, II/2015/DS/1, Completed
WEiTI Działalność statutowa
Languageen angielski
genomics.pdf 301.63 KB
Score (nominal)5
ScoreMinisterial score = 5.0, 01-06-2020, MonographChapterAuthor
Publication indicators WoS Citations = 2; GS Citations = 7.0
Citation count*7 (2020-09-17)
Share Share

Get link to the record

* presented citation count is obtained through Internet information analysis and it is close to the number calculated by the Publish or Perish system.
Are you sure?