proBAMsuite, a Bioinformatics Framework for Genome-Based Representation and Analysis of Proteomics Data

Xiaojing Wang, Robbert J. C. Slebos, Matthew C. Chambers, David L. Tabb, Daniel C. Liebler, Bing Zhang
2015 Molecular & Cellular Proteomics  
To facilitate genome-based representation and analysis of proteomics data, we developed a new bioinformatics framework, proBAMsuite, in which a central component is the protein BAM (proBAM) file format for organizing peptide spectrum matches (PSMs) 1 within the context of the genome. proBAMsuite also includes two R packages, pro-BAMr and proBAMtools, for generating and analyzing pro-BAM files, respectively. Applying proBAMsuite to three recently published proteomics datasets, we demonstrated
more » ... utility in facilitating efficient genome-based sharing, interpretation, and integration of proteomics data. First, the interpretation of proteomics data is significantly enhanced with the rich genomic annotation information. Second, PSMs can be easily reannotated using user-specified gene annotation schemes and assembled into both protein and gene identifications. Third, using the genome as a common reference, proBAMsuite facilitates seamless proteomics and proteogenomics data integration. Finally, proBAM files can be readily visualized in genome browsers and thus bring proteomics data analysis to a general audience beyond the proteomics community. Results from this study establish proBAMsuite as a useful bioinformatics framework for proteomics and proteogenomics research. Molecular & Cellular
doi:10.1074/mcp.m115.052860 pmid:26657539 pmcid:PMC4813696 fatcat:3vlyqxv55fddtku4pca5fh3xze