We support boolean queries, use +,-,<,>,~,* to alter the weighting of terms
A centralized sequence database and community resource for Tribolium genetics, genomics and developmental biology containing genomic sequence scaffolds mapped to 10 linkage groups, genetic linkage maps, the official gene set, Reference Sequences from NCBI (RefSeq), predicted gene models, ESTs and whole-genome tiling array data representing several developmental stages. The current version of Beetlebase is built on the Tribolium castaneum 3.0 Assembly (Tcas 3.0) released by the Human Genome Sequencing Center at the Baylor College of Medicine. The database is constructed using the upgraded Generic Model Organism Database (GMOD) modules. The genomic data is stored in a PostgreSQL relational database using the Chado schema and visualized as tracks in GBrowse. The genetic map is visualized using the comparative genetic map viewer CMAP. To enhance search capabilities, the BLAST search tool has been integrated with the GMOD tools. Tribolium castaneum is a very sophisticated genetic model organism among higher eukaryotes. As the member of a primitive order of holometabolous insects, Coleoptera, Tribolium is in a key phylogenetic position to understand the genetic innovations that accompanied the evolution of higher forms with more complex development. Coleoptera is also the largest and most species diverse of all eukaryotic orders and Tribolium offers the only genetic model for the profusion of medically and economically important species therein. The genome sequences may be downloaded.
Collection of data of protein sequence and functional information. Resource for protein sequence and annotation data. Consortium for preservation of the UniProt databases: UniProt Knowledgebase (UniProtKB), UniProt Reference Clusters (UniRef), and UniProt Archive (UniParc), UniProt Proteomes. Collaboration between European Bioinformatics Institute (EMBL-EBI), SIB Swiss Institute of Bioinformatics and Protein Information Resource. Swiss-Prot is a curated subset of UniProtKB.
Curated, open-source, integrated data resource for comparative functional genomics in crops and model plant species to facilitate the study of cross-species comparisons using information generated from projects supported by public funds. It currently hosts annotated whole genomes in over two dozen plant species and partial assemblies for almost a dozen wild rice species in the Ensembl browser, genetic and physical maps with genes, ESTs and QTLs locations, genetic diversity data sets, structure-function analysis of proteins, plant pathways databases (BioCyc and Plant Reactome platforms), and descriptions of phenotypic traits and mutations. The web-based displays for phenotypes include the Genes and Quantitative Trait Loci (QTL) modules. Sequence based relationships are displayed in the Genomes module using the genome browser adapted from Ensembl, in the Maps module using the comparative map viewer (CMap) from GMOD, and in the Proteins module displays. BLAST is used to search for similar sequences. Literature supporting all the above data is organized in the Literature database. In addition, Gramene now hosts a variety of web services including a Distributed Annotation Server (DAS), BLAST and a public MySQL database. Twice a year, Gramene releases a major build of the database and makes interim releases to correct errors or to make important updates to software and/or data. Additionally you can access Gramene through an FTP site.
Database to catalog experimentally determined interactions between proteins combining information from a variety of sources to create a single, consistent set of protein-protein interactions that can be downloaded in a variety of formats. The data were curated, both, manually and also automatically using computational approaches that utilize the the knowledge about the protein-protein interaction networks extracted from the most reliable, core subset of the DIP data. Because the reliability of experimental evidence varies widely, methods of quality assessment have been developed and utilized to identify the most reliable subset of the interactions. This CORE set can be used as a reference when evaluating the reliability of high-throughput protein-protein interaction data sets, for development of prediction methods, as well as in the studies of the properties of protein interaction networks. Tools are available to analyze, visualize and integrate user's own experimental data with the information about protein-protein interactions available in the DIP database. The DIP database lists protein pairs that are known to interact with each other. By interact they mean that two amino acid chains were experimentally identified to bind to each other. The database lists such pairs to aid those studying a particular protein-protein interaction but also those investigating entire regulatory and signaling pathways as well as those studying the organization and complexity of the protein interaction network at the cellular level. Registration is required to gain access to most of the DIP features. Registration is free to the members of the academic community. Trial accounts for the commercial users are also available.
Collection of pathways and pathway annotations. The core unit of the Reactome data model is the reaction. Entities (nucleic acids, proteins, complexes and small molecules) participating in reactions form a network of biological interactions and are grouped into pathways (signaling, innate and acquired immune function, transcriptional regulation, translation, apoptosis and classical intermediary metabolism) . Provides website to navigate pathway knowledge and a suite of data analysis tools to support the pathway-based analysis of complex experimental and computational data sets.
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 11, 2023. Web tool for an organized view of the transcriptome. Collection of the computationally identified transcripts from the same locus. Information on protein similarities, gene expression, cDNA clones, and genomic location. System for automatically partitioning GenBank sequences into a non redundant set of gene oriented clusters.
Professionally curated repository for genetics, genomics and related data resources for soybean that contains the most current genetic, physical and genomic sequence maps integrated with qualitative and quantitative traits. SoyBase includes annotated Williams 82 genomic sequence and associated data mining tools. The genetic and sequence views of the soybean chromosomes and the extensive data on traits and phenotypes are extensively interlinked. This allows entry to the database using almost any kind of available information, such as genetic map symbols, soybean gene names or phenotypic traits. The repository maintains controlled vocabularies for soybean growth, development, and traits that are linked to more general plant ontologies. Contributions to SoyBase or the Breeder''s Toolbox are welcome.
Public registry of antibodies with unique identifiers for commercial and non-commercial antibody reagents to give researchers a way to universally identify antibodies used in publications. The registry contains antibody product information organized according to genes, species, reagent types (antibodies, recombinant proteins, ELISA, siRNA, cDNA clones). Data is provided in many formats so that authors of biological papers, text mining tools and funding agencies can quickly and accurately identify the antibody reagents they and their colleagues used. The Antibody Registry allows any user to submit a new antibody or set of antibodies to the registry via a web form, or via a spreadsheet upload.
International collaboration producing an extensive public catalog of human genetic variation, including SNPs and structural variants, and their haplotype contexts, in an effort to provide a foundation for investigating the relationship between genotype and phenotype. The genomes of about 2500 unidentified people from about 25 populations around the world were sequenced using next-generation sequencing technologies. Redundant sequencing on various platforms and by different groups of scientists of the same samples can be compared. The results of the study are freely and publicly accessible to researchers worldwide. The consortium identified the following populations whose DNA will be sequenced: Yoruba in Ibadan, Nigeria; Japanese in Tokyo; Chinese in Beijing; Utah residents with ancestry from northern and western Europe; Luhya in Webuye, Kenya; Maasai in Kinyawa, Kenya; Toscani in Italy; Gujarati Indians in Houston; Chinese in metropolitan Denver; people of Mexican ancestry in Los Angeles; and people of African ancestry in the southwestern United States. The goal Project is to find most genetic variants that have frequencies of at least 1% in the populations studied. Sequencing is still too expensive to deeply sequence the many samples being studied for this project. However, any particular region of the genome generally contains a limited number of haplotypes. Data can be combined across many samples to allow efficient detection of most of the variants in a region. The Project currently plans to sequence each sample to about 4X coverage; at this depth sequencing cannot provide the complete genotype of each sample, but should allow the detection of most variants with frequencies as low as 1%. Combining the data from 2500 samples should allow highly accurate estimation (imputation) of the variants and genotypes for each sample that were not seen directly by the light sequencing. All samples from the 1000 genomes are available as lymphoblastoid cell lines (LCLs) and LCL derived DNA from the Coriell Cell Repository as part of the NHGRI Catalog. The sequence and alignment data generated by the 1000genomes project is made available as quickly as possible via their mirrored ftp sites. ftp://ftp.1000genomes.ebi.ac.uk ftp://ftp-trace.ncbi.nlm.nih.gov/1000genomes
Collection of legacy literature in biodiversity assembled by an international consortium of natural history and botanical libraries. It also serves as the foundational literature component of the Encyclopedia of Life. Browse by author, title, subject, collection, map, year, language, and contributor. Taxonomic search using UBio. Also supports data export and a variety of machine interfaces.
Archive for storing and sharing digital data (and accompanying documentation) generated or collected through qualitative and multi method research in social sciences. QDR provides data management consulting services and actively curates all data projects, maintaining value and usefulness of data over time, and ensuring their availability and findability for re-use.
Database for genetic, genomic, phenotype, and disease data generated from rat research. Centralized database that collects, manages, and distributes data generated from rat genetic and genomic research and makes these data available to scientific community. Curation of mapped positions for quantitative trait loci, known mutations and other phenotypic data is provided. Facilitates investigators research efforts by providing tools to search, mine, and analyze this data. Strain reports include description of strain origin, disease, phenotype, genetics, immunology, behavior with links to related genes, QTLs, sub-strains, and strain sources.
Repository of mathematical models of biological and biomedical systems. Hosts selection of existing literature based physiologically and pharmaceutically relevant mechanistic models in standard formats. Features programmatic access via Web Services. Each model is curated to verify that it corresponds to reference publication and gives proper numerical results. Curators also annotate components of models with terms from controlled vocabularies and links to other relevant data resources allowing users to search accurately for models they need. Models can be retrieved in SBML format and import/export facilities are being developed to extend spectrum of formats supported by resource.
Database providing access to information about transmembrane proteins that exist under different conformations, with three primary subfamilies: the cys-loop superfamily, the ATP gated channels superfamily, and the glutamate activated cationic channels superfamily. Due to the lack of evolutionary relationship, these three superfamilies are treated separately. It currently contains 554 entries of ligand-activated ion channel subunits. In this database one may find: the nucleic and proteic sequences of the subunits. Multiple sequence alignments can be generated, and some phylogenetic studies of the superfamilies are provided. Additionally, the atomic coordinates of subunits, or portion of subunits, are provided when available. Redundancy is kept to a minimum, i.e. one entry per gene. Each entry in the database has been manually constructed and checked by a researcher of the field in order to reduce the inaccuracies to a minimum. NOTE: This database is not actively maintained anymore. People should not consider it as an up-to-date trustable resource. For any new work, they should consider using alternative sources, such as UniProt, Ensembl, Protein Databank etc.
Database that provides access to the current and comprehensive 16S rRNA gene sequence alignment for browsing, blasting, probing, and downloading. The data and tools can assist the researcher in choosing phylogenetically specific probes, interpreting microarray results, and aligning/annotating novel sequences. The 16S rRNA gene database provides chimera screening, standard alignment, and taxonomic classification using multiple published taxonomies. ARB users can use Greengenes to update local databases., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Data archive of more than 500,000 files of research in the social sciences, hosting 16 specialized collections of data in education, aging, criminal justice, substance abuse, terrorism, and other fields. ICPSR comprises a consortium of about 700 academic institutions and research organizations providing training in data access, curation, and methods of analysis for the social science research community. ICPSR welcomes and encourages deposits of digital data. ICPSR's educational activities include the Summer Program in Quantitative Methods of Social Research external link, a comprehensive curriculum of intensive courses in research design, statistics, data analysis, and social methodology. ICPSR also leads several initiatives that encourage use of data in teaching, particularly for undergraduate instruction. ICPSR-sponsored research focuses on the emerging challenges of digital curation and data science. ICPSR researchers also examine substantive issues related to our collections, with an emphasis on historical demography and the environment.
Collection of curated, non-redundant genomic DNA, transcript RNA, and protein sequences produced by NCBI. Provides a reference for genome annotation, gene identification and characterization, mutation and polymorphism analysis, expression studies, and comparative analyses. Accessed through the Nucleotide and Protein databases.
The National Institute of Mental Health Data Archive (NDA) makes available human subjects data collected from hundreds of research projects across many scientific domains. Research data repository for data sharing and collaboration among investigators. Used to accelerate scientific discovery through data sharing across all of mental health and other research communities, data harmonization and reporting of research results. Infrastructure created by National Database for Autism Research (NDAR), Research Domain Criteria Database (RDoCdb), National Database for Clinical Trials related to Mental Illness (NDCT), and NIH Pediatric MRI Repository (PedsMRI).
THIS RESOURCE IS NO LONGER IN SERVICE, documented May 10, 2017. A pilot effort that has developed a centralized, web-based biospecimen locator that presents biospecimens collected and stored at participating Arizona hospitals and biospecimen banks, which are available for acquisition and use by researchers. Researchers may use this site to browse, search and request biospecimens to use in qualified studies. The development of the ABL was guided by the Arizona Biospecimen Consortium (ABC), a consortium of hospitals and medical centers in the Phoenix area, and is now being piloted by this Consortium under the direction of ABRC. You may browse by type (cells, fluid, molecular, tissue) or disease. Common data elements decided by the ABC Standards Committee, based on data elements on the National Cancer Institute''s (NCI''s) Common Biorepository Model (CBM), are displayed. These describe the minimum set of data elements that the NCI determined were most important for a researcher to see about a biospecimen. The ABL currently does not display information on whether or not clinical data is available to accompany the biospecimens. However, a requester has the ability to solicit clinical data in the request. Once a request is approved, the biospecimen provider will contact the requester to discuss the request (and the requester''s questions) before finalizing the invoice and shipment. The ABL is available to the public to browse. In order to request biospecimens from the ABL, the researcher will be required to submit the requested required information. Upon submission of the information, shipment of the requested biospecimen(s) will be dependent on the scientific and institutional review approval. Account required. Registration is open to everyone., documented on August 1, 2015. Consortium that aims to facilitate interdisciplinary collaborations to advance the understanding of pancreatic islet development and function, with the goal of developing innovative therapies to correct the loss of beta cell mass in diabetes, including cell reprogramming, regeneration and replacement. They are responsible for collaboratively generating the necessary reagents, mouse strains, antibodies, assays, protocols, technologies and validation assays that are beyond the scope of any single research effort. The scientific goals for the BCBC are to: * Use cues from pancreatic development to directly differentiate pancreatic beta cells and islets from stem / progenitor cells for use in cell-replacement therapies for diabetes, * Determine how to stimulate beta cell regeneration in the adult pancreas as a basis for improving beta cell mass in diabetic patients, * Determine how to reprogram progenitor / adult cells into pancreatic beta-cells both in-vitro and in-vivo as a mean for developing cell-replacement therapies for diabetes, and * Investigate the progression of human type-1 diabetes using patient-derived cells and tissues transplanted in humanized mouse models. Many of the BCBC investigator-initiated projects involve reagent-generating activities that will benefit the larger scientific community. The combination of programs and activities should accelerate the pace of major new discoveries and progress within the field of beta cell biology.
International database for laboratory mouse. Data offered by The Jackson Laboratory includes information on integrated genetic, genomic, and biological data. MGI creates and maintains integrated representation of mouse genetic, genomic, expression, and phenotype data and develops reference data set and consensus data views, synthesizes comparative genomic data between mouse and other mammals, maintains set of links and collaborations with other bioinformatics resources, develops and supports analysis and data submission tools, and provides technical support for database users. Projects contributing to this resource are: Mouse Genome Database (MGD) Project, Gene Expression Database (GXD) Project, Mouse Tumor Biology (MTB) Database Project, Gene Ontology (GO) Project at MGI, and MouseCyc Project at MGI.