Search CORE

5 research outputs found

Big Data Proteogenomics and High Performance Computing: Challenges and Opportunities

Author: Saeed Fahad
Publication venue: ScholarWorks at WMU
Publication date: 01/10/2015
Field of study

Proteogenomics is an emerging field of systems biology research at the intersection of proteomics and genomics. Two high-throughput technologies, Mass Spectrometry (MS) for proteomics and Next Generation Sequencing (NGS) machines for genomics are required to conduct proteogenomics studies. Independently both MS and NGS technologies are inflicted with data deluge which creates problems of storage, transfer, analysis and visualization. Integrating these big data sets (NGS+MS) for proteogenomics studies compounds all of the associated computational problems. Existing sequential algorithms for these proteogenomics datasets analysis are inadequate for big data and high performance computing (HPC) solutions are almost non-existent. The purpose of this paper is to introduce the big data problem of proteogenomics and the associated challenges in analyzing, storing and transferring these data sets. Further, opportunities for high performance computing research community are identified and possible future directions are discussed

ScholarWorks at WMU

Big Data Proteogenomics and High Performance Computing: Challenges and Opportunities

Author: Saeed Fahad
Publication venue: ScholarWorks at WMU
Publication date: 01/10/2015
Field of study

Adam Mickiewicz University Repository

Repozytorium Uniwersytetu im. Adama Mickiewicza (AMUR)

ScholarWorks at WMU

A Scalable Parallel Approach for Peptide Identification from Large-Scale Mass Spectrometry Data

Author
Publication venue: 'Institute of Electrical and Electronics Engineers (IEEE)'
Publication date
Field of study

Crossref