We support boolean queries, use +,-,<,>,~,* to alter the weighting of terms
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on Sep 18, 2018. Open-source, web and programmatically accessible microarray data management system. caArray guides the annotation and exchange of array data using a federated model of local installations whose results are shareable across the cancer Biomedical Informatics Grid (caBIG). caArray furthers translational cancer research through acquisition, dissemination and aggregation of semantically interoperable array data to support subsequent analysis by tools and services on and off the Grid. As array technology advances and matures, caArray will extend its logical library of assay management.
A database that focuses on experimentally verified protein-protein interactions mined from the scientific literature by expert curators. The curated data can be analyzed in the context of the high throughput data and viewed graphically with the MINT Viewer. This collection of molecular interaction databases can be used to search for, analyze and graphically display molecular interaction networks and pathways from a wide variety of species. MINT is comprised of separate database components. HomoMINT, is an inferred human protein interatction database. Domino, is database of domain peptide interactions. VirusMINT explores the interactions of viral proteins with human proteins. The MINT connect viewer allows you to enter a list of proteins (e.g. proteins in a pathway) to retrieve, display and download a network with all the interactions connecting them.
Public depository that collects, annotates, archives, and disseminates important spectral and quantitative data derived from nuclear magnetic resonance spectroscopic investigations of biological macromolecules and metabolites. Provides reference information and maintains a collection of NMR pulse sequences and computer software for biomolecular NMR.
Gene expression data and maps of mouse central nervous system. Gene expression atlas of developing adult central nervous system in mouse, using in situ hybridization and transgenic mouse techniques. Collection of pictorial gene expression maps of brain and spinal cord of mouse. Provides tools to catalog, map, and electrophysiologically record individual cells. Application of Cre recombinase technologies allows for cell-specific gene manipulation. Transgenic mice created by this project are available to scientific community.
Institute to advance genomics in support of the DOE missions related to clean energy generation and environmental characterization and cleanup. Supported by the DOE Office of Science, the DOE JGI unites the expertise at Lawrence Berkeley National Laboratory, Lawrence Livermore National Laboratory, and the HudsonAlpha Institute for Biotechnology. The facility provides integrated high-throughput sequencing and computational analysis that enable systems-based scientific approaches to these challenges.
Database for a curated classification and nomenclature that contains the names of all organisms that are represented in the public sequence databases with at least one nucleotide or protein sequence. Data provided encompasses archaea, bacteria, eukaryota, viroids and viruses. The NCBI taxonomy database is not a primary source for taxonomic or phylogenetic information. Furthermore, the database does not follow a single taxonomic treatise but rather attempts to incorporate phylogenetic and taxonomic knowledge from a variety of sources, including the published literature, web-based databases, and the advice of sequence submitters and outside taxonomy experts. Consequently, the NCBI taxonomy database is not a phylogenetic or taxonomic authority and should not be cited as such.
Database of three-dimensional structures of macromolecules that allows the user to retrieve structures for specific molecule types as well as structures for genes and proteins of interest. Three main databases comprise Structure-The Molecular Modeling Database; Conserved Domains and Protein Classification; and the BioSystems Database. Structure also links to the PubChem databases to connect biological activity data to the macromolecular structures. Users can locate structural templates for proteins and interactively view structures and sequence data to closely examine sequence-structure relationships. * Macromolecular structures: The three-dimensional structures of biomolecules provide a wealth of information on their biological function and evolutionary relationships. The Molecular Modeling Database (MMDB), as part of the Entrez system, facilitates access to structure data by connecting them with associated literature, protein and nucleic acid sequences, chemicals, biomolecular interactions, and more. It is possible, for example, to find 3D structures for homologs of a protein of interest by following the Related Structure link in an Entrez Protein sequence record. * Conserved domains and protein classification: Conserved domains are functional units within a protein that act as building blocks in molecular evolution and recombine in various arrangements to make proteins with different functions. The Conserved Domain Database (CDD) brings together several collections of multiple sequence alignments representing conserved domains, in addition to NCBI-curated domains that use 3D-structure information explicitly to define domain boundaries and provide insights into sequence/structure/function relationships. * Small molecules and their biological activity: The PubChem project provides information on the biological activities of small molecules and is a component of NIH''''s Molecular Libraries Roadmap Initiative. PubChem includes three databases: PCSubstance, PCBioAssay, and PCCompound. The PubChem data are linked to other data types (illustrated example) in the Entrez system, making it possible, for example, to retrieve information about a compound and then Link to its biological activity data, retrieve 3D protein structures bound to the compound and interactively view their active sites, and find biosystems that include the compound as a component. * Biological Systems: A biosystem, or biological system, is a group of molecules that interact directly or indirectly, where the grouping is relevant to the characterization of living matter. The NCBI BioSystems Database provides centralized access to biological pathways from several source databases and connects the biosystem records with associated literature, molecular, and chemical data throughout the Entrez system. BioSystem records list and categorize components (illustrated example), such as the genes, proteins, and small molecules involved in a biological system. The companion FLink icon FLink tool, in turn, allows you to input a list of proteins, genes, or small molecules and retrieve a ranked list of biosystems.
Public registry of nucleic acid reagents designed for use in a wide variety of biomedical research applications including genotyping, gene expression studies, SNP discovery, genome mapping, and gene silencing. Probe records contain information on reagent distributors, probe effectiveness, and computed sequence similarities. The database is constantly updated, with over 11,000,000 probes available. Users may deposit their data into NCBI Probe Database.
Multidisciplinary data repository for a consortium of universities in the Netherlands housing over datasets with a focus on scientific and technical data. Most data were produced by Dutch researchers including datasets from doctoral research. Users can deposit up to 1G by completing an upload form. Collection development foci include applied sciences, biomedical technology, earth sciences, and technology and construction. 4TU.Datacentrum is a collaboration of the libraries of the three leading technical universities - Delft University of Technology, Eindhoven University of Technology and the University of Twente.
Database providing integrated access to genome sequence, expression data and literature curation for Tuberculosis (TB) that houses genome assemblies for numerous strains of Mycobacterium tuberculosis (MTB) as well assemblies for over 20 strains related to MTB and useful for comparative analysis. TBDB stores pre- and post-publication gene-expression data from M. tuberculosis and its close relatives, including over 3000 MTB microarrays, 95 RT-PCR datasets, 2700 microarrays for human and mouse TB related experiments, and 260 arrays for Streptomyces coelicolor. (July 2010) To enable wide use of these data, TBDB provides a suite of tools for searching, browsing, analyzing, and downloading the data.
Functional genomics data repository supporting MIAME-compliant data submissions. Tools are provided to help users query and download experiments and curated gene expression profiles. These data include microarray-based experiments measuring the abundance of mRNA, genomic DNA, and protein molecules, as well as non-array-based technologies such as serial analysis of gene expression (SAGE) and mass spectrometry proteomic technology. Array- and sequence-based data are accepted.
A large multidisciplinary world leader in genomic research with locations in Rockville, Maryland and San Diego, California. It was formed through the merger of several affiliated and legacy organizations - The Institute for Genomic Research (TIGR) and The Center for the Advancement of Genomics (TCAG), The J. Craig Venter Science Foundation, The Joint Technology Center, and the Institute for Biological Energy Alternatives (IBEA).
Bioinformatics Resource Center for invertebrate vectors. Provides web-based resources to scientific community conducting basic and applied research on organisms considered potential agents of biowarfare or bioterrorism or causing emerging or re-emerging diseases.
An Internet library offering the general public access to historical collections that exist in digital format including texts, audio, moving images, and software. Additionally it provides archived web pages in their collections, and specialized services for adaptive reading and information access for the blind and other persons with disabilities. Founded in 1996 and located in San Francisco, the Archive has been receiving data donations from Alexa Internet and others. In late 1999, the organization started to grow to include more well-rounded collections.
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 23, 2019 Database was open, publicly accessible platform for DNA and clinical data related to human Major Histocompatibility Complex (MHC). Data from IHWG workshops were provided as well., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
THIS RESOURCE IS NO LONGER IN SERVICE, documented on May 11, 2016. Repository of brain-mapping data (surfaces and volumes; structural and functional data) derived from studies including fMRI and MRI from many laboratories, providing convenient access to a growing body of neuroimaging and related data. WebCaret is an online visualization tool for viewing SumsDB datasets. SumsDB includes: * data on cerebral cortex and cerebellar cortex * individual subject data and population data mapped to atlases * data from FreeSurfer and other brainmapping software besides Caret SumsDB provides multiple levels of data access and security: * Free (public) access (e.g., for data associated with published studies) * Data access restricted to collaborators in different laboratories * Owner-only access for work in progress Data can be downloaded from SumsDB as individual files or as bundles archived for offline visualization and analysis in Caret WebCaret provides online Caret-style visualization while circumventing software and data downloads. It is a server-side application running on a linux cluster at Washington University. WebCaret "scenes" facilitate rapid visualization of complex combinations of data Bi-directional links between online publications and WebCaret/SumsDB provide: * Links from figures in online journal article to corresponding scenes in WebCaret * Links from metadata in WebCaret directly to relevant online publications and figures
Central data repository for nematode biology including complete genomic sequence, gene predictions and orthology assignments from range of related nematodes.Data concerning genetics, genomics and biology of C. elegans and related nematodes. Derived from initial ACeDB database of C. elegans genetic and sequence information, WormBase includes genomic, anatomical and functional information of C. elegans, other Caenorhabditis species and other nematodes. Maintains public FTP site where researchers can find many commonly requested files and datasets, WormBase software and prepackaged databases.
Databases of protein sequences and 3D structures of proteins. Collection of sequences from several sources, including translations from annotated coding regions in GenBank, RefSeq and TPA, as well as records from SwissProt, PIR, PRF, and PDB.
DNA barcode data with an online workbench that supports data validation, annotation, and publication for specimen, distributional, and molecular data. The data platform consists of three main modules, a data portal, a database of barcode clusters, and data collection workbench. The Public Data Portal provides access to all public barcode data which consists of data generated using the Workbench module as well as data mined from other sources. The Barcode Index Number (BIN) system assigns a unique identifier to each sequence cluster of COI, providing an interim taxonomic system for species in the animal kingdom. The workbench module integrates secure databases with analytical tools to provide a private collaborative environment for researchers to collect, analyze, and publish barcode data and ancillary DNA sequences. This platform also provides an annotation framework that supports tagging and commenting on records and their components (i.e. taxonomy, images, and sequences), allowing for community-based validation of barcode data. By providing specialized services, it aids in the assembly of records that meet the standards needed to gain BARCODE designation in the global sequence databases. Because of its web-based delivery and flexible data security model, it is also well positioned to support projects that involve broad research alliances. Public data records include record identifiers, taxonomy, specimen details, collection information and sequence data. Data that has been publicly released through BOLD can be retrieved manually through the BOLD public interface or automatically through BOLD web services. BOLD analytical tools are available for any data set that exists in BOLD (including publicly available data). Analytical tools can be accessed through the BOLD Project Console under the headings Sequences Analysis or Specimen Aggregates. Some examples include Taxon ID Tree, Alignment Viewer, Distribution Maps, and Image Library.
Database of nucleotide sequences from several sources, including GenBank, RefSeq, TPA and PDB. Genome, gene and transcript sequence data provide the foundation for biomedical research and discovery.