Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
Elucidating the complete network of protein-protein interactions, or interactome, is a fundamental goal of the post-genomic era, yet existing interactome maps are far from complete. To increase the throughput and resolution of interactome mapping, methods for protein-protein interaction discovery by co-migration have been introduced. However, accurate identification of interacting protein pairs within the resulting large-scale proteomic datasets is challenging. Consequently, most computational pipelines for co-migration data analysis incorporate external genomic datasets to distinguish interacting from non-interacting protein pairs. The effect of this procedure on interactome mapping is poorly understood. Here, we conduct a rigorous analysis of genomic data integration for interactome recovery across a large number of co-migration datasets, spanning diverse experimental and computational methods. We find that genomic data integration leads to an increase in the functional coherence of the resulting interactome maps, but this comes at the expense of a decrease in power to discover novel interactions. Importantly, putative novel interactions predicted by genomic data integration are no more likely to later be experimentally discovered than those predicted from co-migration data alone. Our results reveal a widespread and unappreciated limitation in a methodology that has been widely used to map the interactome of humans and model organisms.
Pubmed ID: 30332399
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
A database that focuses on experimentally verified protein-protein interactions mined from the scientific literature by expert curators. The curated data can be analyzed in the context of the high throughput data and viewed graphically with the MINT Viewer. This collection of molecular interaction databases can be used to search for, analyze and graphically display molecular interaction networks and pathways from a wide variety of species. MINT is comprised of separate database components. HomoMINT, is an inferred human protein interatction database. Domino, is database of domain peptide interactions. VirusMINT explores the interactions of viral proteins with human proteins. The MINT connect viewer allows you to enter a list of proteins (e.g. proteins in a pathway) to retrieve, display and download a network with all the interactions connecting them.
View all literature mentionsFreely available database focused on interactions established by extracellular proteins and polysaccharides, taking into account the multimeric nature of the extracellular proteins (e.g. collagens, laminins and thrombospondins are multimers). MatrixDB is an active member of the International Molecular Exchange (IMEx) consortium and has adopted the PSI-MI standards for annotating and exchanging interaction data. It includes interaction data extracted from the literature by manual curation, and offers access to relevant data involving extracellular proteins provided by the IMEx partner databases through the PSICQUIC webservice, as well as data from the Human Protein Reference Database. The database reports mammalian protein-protein and protein-carbohydrate interactions involving extracellular molecules. Interactions with lipids and cations are also reported. MatrixDB is focused on mammalian interactions, but aims to integrate interaction datasets of model organisms when available. MatrixDB provides direct links to databases recapitulating mutations in genes encoding extracellular proteins, to UniGene and to the Human Protein Atlas that shows expression and localization of proteins in a large variety of normal human tissues and cells. MatrixDB allows researchers to perform customized queries and to build tissue- and disease-specific interaction networks that can be visualized and analyzed with Cytoscape or Medusa. Statistics (2013): 2283 extracellular matrix interactions including 2095 protein-protein and 169 protein-glycosaminoglycan interactions.
View all literature mentionsOpen and collaborative platform dedicated to curation of biological pathways. Each pathway has dedicated wiki page, displaying current diagram, description, references, download options, version history, and component gene and protein lists. Database of biological pathways maintained by and for scientific community.
View all literature mentions
Database of manually annotated protein complexes from mammalian organisms. Annotation includes protein complex function, localization, subunit composition, literature references and more. All information is obtained from individual experiments published in scientific articles, but data from high-throughput experiments is excluded.
The majority of protein complexes in CORUM originates from man (65%), followed by mouse (14%) and rat (14%).
Collection of data of protein sequence and functional information. Resource for protein sequence and annotation data. Consortium for preservation of the UniProt databases: UniProt Knowledgebase (UniProtKB), UniProt Reference Clusters (UniRef), and UniProt Archive (UniParc), UniProt Proteomes. Collaboration between European Bioinformatics Institute (EMBL-EBI), SIB Swiss Institute of Bioinformatics and Protein Information Resource. Swiss-Prot is a curated subset of UniProtKB.
View all literature mentionsA database of high-quality protein-protein interactions in different organisms.
View all literature mentionsA manually curated resource of signal transduction pathways in humans. All pathways are freely available for download in BioPAX level 3.0, PSI-MI version 2.5 and SBML version 2.1 formats. The slim pathway models representing only core reactions in each pathway are available at NetSlim. All the NetPath pathway models are also submitted to WikiPathways.
View all literature mentionsA database of protein families, each represented by multiple sequence alignments and hidden Markov models (HMMs). Users can analyze protein sequences for Pfam matches, view Pfam family annotation and alignments, see groups of related families, look at the domain organization of a protein sequence, find the domains on a PDB structure, and query Pfam by keywords. There are two components to Pfam: Pfam-A and Pfam-B. Pfam-A entries are high quality, manually curated families that may automatically generate a supplement using the ADDA database. These automatically generated entries are called Pfam-B. Although of lower quality, Pfam-B families can be useful for identifying functionally conserved regions when no Pfam-A entries are found. Pfam also generates higher-level groupings of related families, known as clans (collections of Pfam-A entries which are related by similarity of sequence, structure or profile-HMM).
View all literature mentionsPublic bibliographic database that provides access to citations for biomedical literature from MEDLINE, life science journals, and online books. Citations may include links to full-text content from PubMed Central and publisher web sites. PubMed citations and abstracts include fields of biomedicine and health, covering portions of life sciences, behavioral sciences, chemical sciences, and bioengineering. Provides access to additional relevant web sites and links to other NCBI molecular biology resources. Publishers of journals can submit their citations to NCBI and then provide access to full-text of articles at journal web sites using LinkOut.
View all literature mentionsDatabase of known and predicted protein interactions. The interactions include direct (physical) and indirect (functional) associations and are derived from four sources: Genomic Context, High-throughput experiments, (Conserved) Coexpression, and previous knowledge. STRING quantitatively integrates interaction data from these sources for a large number of organisms, and transfers information between these organisms where applicable. The database currently covers 5''214''234 proteins from 1133 organisms. (2013)
View all literature mentionsDatabase that represents a centralized platform to visually depict and integrate information pertaining to domain architecture, post-translational modifications, interaction networks and disease association for each protein in the human proteome. All the information in HPRD has been manually extracted from the literature by expert biologists who read, interpret and analyze the published data.
View all literature mentionsCurated protein-protein and genetic interaction repository of raw protein and genetic interactions from major model organism species, with data compiled through comprehensive curation efforts.
View all literature mentionsTHIS RESOURCE IS NO LONGER IN SERVICE, documented on February 1st, 2022. Software application for genetic analysis of classical biometric traits like blood pressure or height that are caused by a combination of polygenic inheritance and complex environmental forces. (entry from Genetic Analysis Software)
View all literature mentionsWeb tool used to generate confidence scored and functionally annotated human protein-protein interaction networks. Users can run a protein query using a UniProt identifier (ID or accession), gene symbol or Entrez gene id, or a network query using set fields.
View all literature mentionsSoftware that archives evidence collected from different sources, then analyzes and presents these data. Its data come from manually curated protein-protein interaction databases that have adhered to the IMEx consortium.
View all literature mentionsCell line HeLa is a Cancer cell line with a species of origin Homo sapiens
View all literature mentionsCell line Jurkat is a Cancer cell line with a species of origin Homo sapiens (Human)
View all literature mentions