Back to Search Start Over

Compression of Next-Generation Sequencing Data and of DNA Digital Files †.

Authors :
Carpentieri, Bruno
Source :
Algorithms. Jun2020, Vol. 13 Issue 6, p151-151. 1p.
Publication Year :
2020

Abstract

The increase in memory and in network traffic used and caused by new sequenced biological data has recently deeply grown. Genomic projects such as HapMap and 1000 Genomes have contributed to the very large rise of databases and network traffic related to genomic data and to the development of new efficient technologies. The large-scale sequencing of samples of DNA has brought new attention and produced new research, and thus the interest in the scientific community for genomic data has greatly increased. In a very short time, researchers have developed hardware tools, analysis software, algorithms, private databases, and infrastructures to support the research in genomics. In this paper, we analyze different approaches for compressing digital files generated by Next-Generation Sequencing tools containing nucleotide sequences, and we discuss and evaluate the compression performance of generic compression algorithms by confronting them with a specific system designed by Jones et al. specifically for genomic file compression: Quip. Moreover, we present a simple but effective technique for the compression of DNA sequences in which we only consider the relevant DNA data and experimentally evaluate its performances. [ABSTRACT FROM AUTHOR]

Details

Language :
English
ISSN :
19994893
Volume :
13
Issue :
6
Database :
Academic Search Index
Journal :
Algorithms
Publication Type :
Academic Journal
Accession number :
144697921
Full Text :
https://doi.org/10.3390/a13060151