Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Using high abundance proteins as guides for fast and effective peptide/protein identification from metaproteomic data

View through CrossRef
Abstract Background A few recent large efforts significantly expanded the collection of human-associated bacterial genomes, which now contains thousands of entities including reference complete/draft genomes and metagenome assembled genomes (MAGs). These genomes provide useful resource for studying the functionality of the human-associated microbiome and their relationship with human health and diseases. One application of these genomes is to provide a universal reference for database search in metaproteomic studies, when matched metagenomic/metatranscriptomic data are unavailable. However, a greater collection of reference genomes may not necessarily result in better peptide/protein identification because the increase of search space often leads to fewer spectrum-peptide matches, not to mention the drastic increase of computation time. Methods Here, we present a new approach that uses two steps to optimize the use of the reference genomes and MAGs as the universal reference for human gut metaproteomic MS/MS data analysis. The first step is to use only the High Abundance Proteins (HAPs) (i.e., ribosomal proteins and elongation factors) for metaproteomic MS/MS database search and, based on the identification results, to derive the taxonomic composition of the underlying microbial community. The second step is to expand the search database by including all proteins from identified abundant species. We call our approach HAPiID (HAPs guided metaproteomics IDentification). Results We tested our approach using human gut metaproteomic datasets from a previous study and compared it to the state-of-the-art reference database search method MetaPro-IQ for metaproteomic identification in studying human gut microbiota. Our results show that our two-steps method not only performed significantly faster but also was able to identify more peptides. We further demonstrated the application of HAPiID to revealing protein profiles of individual human-associated bacterial species, one or a few species at a time, using metaproteomic data. Conclusions The HAP guided profiling approach presents a novel effective way for constructing target database for metaproteomic data analysis. The HAPiID pipeline built upon this approach provides a universal tool for analyzing human gut-associated metaproteomic data.
Title: Using high abundance proteins as guides for fast and effective peptide/protein identification from metaproteomic data
Description:
Abstract Background A few recent large efforts significantly expanded the collection of human-associated bacterial genomes, which now contains thousands of entities including reference complete/draft genomes and metagenome assembled genomes (MAGs).
These genomes provide useful resource for studying the functionality of the human-associated microbiome and their relationship with human health and diseases.
One application of these genomes is to provide a universal reference for database search in metaproteomic studies, when matched metagenomic/metatranscriptomic data are unavailable.
However, a greater collection of reference genomes may not necessarily result in better peptide/protein identification because the increase of search space often leads to fewer spectrum-peptide matches, not to mention the drastic increase of computation time.
Methods Here, we present a new approach that uses two steps to optimize the use of the reference genomes and MAGs as the universal reference for human gut metaproteomic MS/MS data analysis.
The first step is to use only the High Abundance Proteins (HAPs) (i.
e.
, ribosomal proteins and elongation factors) for metaproteomic MS/MS database search and, based on the identification results, to derive the taxonomic composition of the underlying microbial community.
The second step is to expand the search database by including all proteins from identified abundant species.
We call our approach HAPiID (HAPs guided metaproteomics IDentification).
Results We tested our approach using human gut metaproteomic datasets from a previous study and compared it to the state-of-the-art reference database search method MetaPro-IQ for metaproteomic identification in studying human gut microbiota.
Our results show that our two-steps method not only performed significantly faster but also was able to identify more peptides.
We further demonstrated the application of HAPiID to revealing protein profiles of individual human-associated bacterial species, one or a few species at a time, using metaproteomic data.
Conclusions The HAP guided profiling approach presents a novel effective way for constructing target database for metaproteomic data analysis.
The HAPiID pipeline built upon this approach provides a universal tool for analyzing human gut-associated metaproteomic data.

Related Results

Frequency of Common Chromosomal Abnormalities in Patients with Idiopathic Acquired Aplastic Anemia
Frequency of Common Chromosomal Abnormalities in Patients with Idiopathic Acquired Aplastic Anemia
Objective: To determine the frequency of common chromosomal aberrations in local population idiopathic determine the frequency of common chromosomal aberrations in local population...
7 th International Symposium on Enabling Technologies for Life Sciences (ETP)
7 th International Symposium on Enabling Technologies for Life Sciences (ETP)
The seventh in the series of ETP Symposia (see Rapid Communications in Mass Spectrometry 2012, 26 , ...
Metaproteomic profiling of the secretome of a granule-forming Ca . Accumulibacter enrichment
Metaproteomic profiling of the secretome of a granule-forming Ca . Accumulibacter enrichment
ABSTRACT Extracellular proteins are supposed to play crucial roles in the formation and structure of biofilms and aggregates. However, often litt...
Protein-peptide Interaction Representation Learning with Pretrained Language Models
Protein-peptide Interaction Representation Learning with Pretrained Language Models
Abstract Protein-peptide Interactions (PpIs) paly essential roles in diverse cellular processes, yet their systematic identification remains challenging due to the ...
Anemia Is Inversely Associated with Serum C-Peptide Concentrations in Patients with Type 2 Diabetes
Anemia Is Inversely Associated with Serum C-Peptide Concentrations in Patients with Type 2 Diabetes
Results: The aim of the study was to investigate the relationship between anemia and serum C-peptide concentrations in Korean patients with type 2 diabetes. A total of 1,300 subjec...
Endothelial Protein C Receptor
Endothelial Protein C Receptor
IntroductionThe protein C anticoagulant pathway plays a critical role in the negative regulation of the blood clotting response. The pathway is triggered by thrombin, which allows ...
Peptide-Directed Supramolecular Self-Assembly of N-Substituted Perylene Imides
Peptide-Directed Supramolecular Self-Assembly of N-Substituted Perylene Imides
<p>Synthetic peptides offer enormous potential to encode the assembly of molecular electronic components, provided that the complex range of interactions is distilled into si...

Back to Top