Ph.D. Dissertation Defense - Nagakishore Jammula

Event Details

Thursday, November 29, 2018

1:00pm - 3:00pm

Location: 
Room 1212, Klaus

For More Information

Contact:

Event Details

TitleParallel Algorithms for Enabling Fast and Scalable Analysis of High-throughput Sequencing Datasets

Committee:

Dr. Srinivas Aluru, CoC, Chair , Advisor

Dr. Richard Vuduc, CoC

Dr. Moinuddin Qureshi, ECE

Dr. Linda Wills, ECE

Dr. Ada Gavrilovska, CoC

Abstract: 

The objective of this research is to develop parallel algorithms for enabling fast and scalable analysis of large-scale high-throughput sequencing datasets. Genome of an organism consists of one or more long DNA sequences called chromosomes, each a sequence of bases. Depending on the organism, the length of the genome can vary from several thousand bases to several billion bases. Genome sequencing, which involves deciphering the sequence of bases of the genome, is an important tool in genomics research. Sequencing instruments in vogue today can only read short DNA sequences. However, these instruments can read billions of such sequences at a time, and are used to sequence a large number of randomly generated short genomic fragments from the genome. These fragments are a few hundred bases long and are commonly referred to as “reads”. This work specifically tackles three problems associated with high-throughput sequencing datasets: (1) Parallel read error correction for large-scale genomics datasets, (2) Partitioning of large-scale high-throughput sequencing datasets, and (3) Parallel compression of large-scale genomics datasets.

Last revised November 26, 2018