Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
A genome browser specialized in next-generation sequencing data.
Proper citation: GenomeJack (RRID:SCR_012026) Copy
A database of genomic and protein data for Drosophila site-specific transcription factors.
Proper citation: FlyTF.org (RRID:SCR_004123) Copy
http://www.ncbi.nlm.nih.gov/mapview/map_search.cgi?taxid=7165
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 11, 2023. A database for the Anopheles gambiae str. PEST genome that was sequenced using a whole genome shotgun approach. The database aims to contribute to the understanding of mosquito genome structure and organization and will assist the development of malaria control strategies and improved anti-malarial drugs and vaccines. Sequences were generated and assembled into contigs for submission to GenBank.
Proper citation: Anopheles gambiae (African malaria mosquito) genome view (RRID:SCR_004402) Copy
Database of the international consortium working together to mutate all protein-coding genes in the mouse using a combination of gene trapping and gene targeting in C57BL/6 mouse embryonic stem (ES) cells. Detailed information on targeted genes is available. The IKMC includes the following programs: * Knockout Mouse Project (KOMP) (USA) ** CSD, a collaborative team at the Children''''s Hospital Oakland Research Institute (CHORI), the Wellcome Trust Sanger Institute and the University of California at Davis School of Veterinary Medicine , led by Pieter deJong, Ph.D., CHORI, along with K. C. Kent Lloyd, D.V.M., Ph.D., UC Davis; and Allan Bradley, Ph.D. FRS, and William Skarnes, Ph.D., at the Wellcome Trust Sanger Institute. ** Regeneron, a team at the VelociGene division of Regeneron Pharmaceuticals, Inc., led by David Valenzuela, Ph.D. and George D. Yancopoulos, M.D., Ph.D. * European Conditional Mouse Mutagenesis Program (EUCOMM) (Europe) * North American Conditional Mouse Mutagenesis Project (NorCOMM) (Canada) * Texas A&M Institute for Genomic Medicine (TIGM) (USA) Products (vectors, mice, ES cell lines) may be ordered from the above programs.
Proper citation: International Knockout Mouse Consortium (RRID:SCR_005574) Copy
http://swissregulon.unibas.ch/fcgi/sr/swissregulon
A database of genome-wide annotations of regulatory sites. The predictions are based on Bayesian probabilistic analysis of a combination of input information including: * Experimentally determined binding sites reported in the literature. * Known sequence-specificities of transcription factors. * ChIP-chip and ChIP-seq data. * Alignments of orthologous non-coding regions. Predictions were made using the PhyloGibbs, MotEvo, IRUS and ISMARA algorithms developed in their group, depending on the data available for each organism. Annotations can be viewed in a Gbrowse genome browser and can also be downloaded in flat file format.
Proper citation: SwissRegulon (RRID:SCR_005333) Copy
A publicly available database of Transposed elements (TEs) which are located within protein-coding genes of 7 organisms: human, mouse, chicken, zebrafish, fruilt fly, nematode and sea squirt. Using TranspoGene the user can learn about the many aspects of the effect these TEs have on their hosting genes, such as: exonization events (including alternative splicing-related data), insertion of TEs into introns, exons, and promoters, specific location of the TE over the gene, evolutionary divergence of the TE from its consensus sequence and involvement in diseases. TranspoGene database is quickly searchable through its website, enables many kinds of searches and is available for download. TranspoGene contains information regarding specific type and family of the TEs, genomic and mRNA location, sequence, supporting transcript accession and alignment to the TE consensus sequence. The database also contains host gene specific data: gene name, genomic location, Swiss-Prot and RefSeq accessions, diseases associated with the gene and splicing pattern. The TranspoGene and microTranspoGene databases can be used by researchers interested in the effect of TE insertion on the eukaryotic transcriptome.
Proper citation: TranspoGene (RRID:SCR_005634) Copy
A knowledgebase of Biochemically, Genetically and Genomically structured genome-scale metabolic network reconstructions. BiGG integrates several published genome-scale metabolic networks into one resource with standard nomenclature which allows components to be compared across different organisms. BiGG can be used to browse model content, visualize metabolic pathway maps, and export SBML files of the models for further analysis by external software packages. Users may follow links from BiGG to several external databases to obtain additional information on genes, proteins, reactions, metabolites and citations of interest.
Proper citation: BiGG Database (RRID:SCR_005809) Copy
http://h-invitational.jp/varygene/
It consists of a Genome Browser, an LD Search System, and the VaryGene 2 system. The Generic Genome Browser is a combination of database and interactive Web page for manipulating and displaying annotations on genomes, while LDSearchSystem is a search system for linkage disequilibrium (LD) bins. VaryGene 2 is a system to search, display, and download our research results on human polymorphism based on publicly available data and annotations of transcripts presented by H-InvDB. VaryGene 2 provides information about single nucleotide polymorphisms (SNPs), deletion-insertion polymorphisms (DIPs), short tandem repeats (STRs), single amino acid repeats (SARs), structural variation (or copy number variations: CNVs), and their relations to the genome, transcripts, and functional domains. Users can search by polymorphisms, transcripts, STRs/SARs, and CNVs.
Proper citation: VarySysDB (RRID:SCR_005880) Copy
ProPortal is a database containing genomic, metagenomic, transcriptomic and field data for the marine cyanobacterium Prochlorococcus. Our goal is to provide a source of cross-referenced data across multiple scales of biological organization--from the genome to the ecosystem--embracing the full diversity of ecotypic variation within this microbial taxon, its sister group, Synechococcus and phage that infect them. The site currently contains the genomes of 13 Prochlorococcus strains, 11 Synechococcus strains and 28 cyanophage strains that infect one or both groups. Cyanobacterial and cyanophage genes are clustered into orthologous groups that can be accessed by keyword search or through a genome browser. Users can also identify orthologous gene clusters shared by cyanobacterial and cyanophage genomes. Gene expression data for Prochlorococcus ecotypes MED4 and MIT9313 allow users to identify genes that are up or downregulated in response to environmental stressors. In addition, the transcriptome in synchronized cells grown on a 24-h light-dark cycle reveals the choreography of gene expression in cells in a ''natural'' state. Metagenomic sequences from the Global Ocean Survey from Prochlorococcus, Synechococcus and phage genomes are archived so users can examine the differences between populations from diverse habitats. Finally, an example of cyanobacterial population data from the field is included.
Proper citation: ProPortal (RRID:SCR_006112) Copy
This database presents the entire DNA sequence of the first diploid genome sequence of a Han Chinese, a representative of Asian population. The genome, named as YH, represents the start of YanHuang Project, which aims to sequence 100 Chinese individuals in 3 years. It was assembled based on 3.3 billion reads (117.7Gbp raw data) generated by Illumina Genome Analyzer. In total of 102.9Gbp nucleotides were mapped onto the NCBI human reference genome (Build 36) by self-developed software SOAP (Short Oligonucleotide Alignment Program), and 3.07 million SNPs were identified. The personal genome data is illustrated in a MapView, which is powered by GBrowse. A new module was developed to browse large-scale short reads alignment. This module enabled users track detailed divergences between consensus and sequencing reads. In total of 53,643 HGMD recorders were used to screen YH SNPs to retrieve phenotype related information, to superficially explain the donor's genome. Blast service to align query sequences against YH genome consensus was also provided.
Proper citation: YanHuang Project (RRID:SCR_006077) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 15, 2013. TRIPLES provides full public access to the data and reagents generated from ongoing functional analysis of the yeast genome. Using a novel transposon-tagging approach, we have analyzed disruption phenotypes, gene expression, and protein localization on a genome-wide scale in Saccharomyces. The data generated from this study may be accessed through our database, TRIPLES ; additionally, all reagents generated in this study are freely available from on-line order forms (linked to TRIPLES as well). multipurpose, mini-transposon, mutant alleles, phenotypes, protein localization, gene expression, Saccharomyces cerevisiae, Web-accessible database, transposon-mutagenized yeast strains, downloaded, tab-delimited, text file, protein localization data, fluorescent micrographs, staining patterns, indirect immunofluorescence analysis of indicated epitope-tagged proteins, subcellular localization of the yeast proteome, visual library, Nucleic Acid Sequence Data Library (GenBank), clone report, graphic map, transposon insertions (represented as flags)
Proper citation: TRIPLES- a database of TRansposon-Insertion Phenotypes Localization and Expression in Saccharomyces (RRID:SCR_005714) Copy
http://igdb.nsclc.ibms.sinica.edu.tw/
IGDB.NSCLC database is aiming to facilitate and prioritize identified lung cancer genes and microRNAs for pathological and mechanistic studies of lung tumorigenesis and for developing new strategies for clinical interventions. We integrated and curated various lung cancer genomic datasets to present # lung cancer genes with somatic mutations, experimental supports and statistic significance in association with clinicopathological features; # genomic alterations with copy number alterations (CNA) detected by high density SNP arrays, gain or loss regions detected by arrayed comparative genome hybridization (aCGH), and loss of heterozygosity (LOH) detected by microsatellite markers; # aberrant expression of genes and microRNAs detected by various microarrays. IGDB.NSCLC database provides user friendly interfaces and searching functions to display multiple layers of evidence for detecting lung cancer target genes and microRNAs, especially emphasizing on concordant alterations: # genes with altered expression located in the CNA regions; # microRNAs with altered expression located in the CNA regions; # somatic mutation genes located in the CNA regions; and # genes associated with clinicopathological features located in the CNA regions. These concordant altered genes and miRNAs should be prioritized for further basic and clinical studies.
Proper citation: IGDB.NSCLC (RRID:SCR_006048) Copy
A comprehensive biochemical knowledge-base on human metabolism, this community-driven, consensus metabolic reconstruction integrates metabolic information from five different resources: * Recon 1, a global human metabolic reconstruction (Duarte et al, PNAS, 104(6), 1777-1782, 2007) * EHMN, Edinburgh Human Metabolic Network (Hao et al., BMC Bioinformatics 11, 393, 2010) * HepatoNet1, a liver metabolic reconstruction (Gille et al., Molecular Systems Biology 6, 411, 2010), * Ac/FAO module, an acylcarnitine/fatty acid oxidation module (Sahoo et al., Molecular bioSystems 8, 2545-2558, 2012), * a human small intestinal enterocytes reconstruction (Sahoo and Thiele, submitted). Additionally, more than 370 transport and exchange reactions were added, based on a literature review. Recon 2 is fully semantically annotated (Le Nov��re, N. et al. Nat Biotechnol 23, 1509-1515, 2005) with references to persistent and publicly available chemical and gene databases, unambiguously identifying its components and increasing its applicability for third-party users. Here you can explore the content of the reconstruction by searching/browsing metabolites and reactions. Recon 2 predictive model is available in the Systems Biology Markup Language format.
Proper citation: Recon x (RRID:SCR_006345) Copy
http://www.informatics.jax.org
International database for laboratory mouse. Data offered by The Jackson Laboratory includes information on integrated genetic, genomic, and biological data. MGI creates and maintains integrated representation of mouse genetic, genomic, expression, and phenotype data and develops reference data set and consensus data views, synthesizes comparative genomic data between mouse and other mammals, maintains set of links and collaborations with other bioinformatics resources, develops and supports analysis and data submission tools, and provides technical support for database users. Projects contributing to this resource are: Mouse Genome Database (MGD) Project, Gene Expression Database (GXD) Project, Mouse Tumor Biology (MTB) Database Project, Gene Ontology (GO) Project at MGI, and MouseCyc Project at MGI.
Proper citation: Mouse Genome Informatics (MGI) (RRID:SCR_006460) Copy
http://www.snpedia.com/index.php/SNPedia
Wiki investigating human genetics including information about the effects of variations in DNA, citing peer-reviewed scientific publications. It is used by Promethease to analyze and help explain your DNA. It is based on a wiki model in order to foster communication about genetic variation and to allow interested community members to help it evolve to become ever more relevant. As the cost of genotyping (and especially of fully determining your own genomic sequence) continues to drop, we''''ll all want to know more - a lot more - about the meaning of these DNA variations and SNPedia will be here to help. SNPedia has been launched to help realize the potential of the Human Genome Project to connect to our daily lives and well-being. For more information see the Wikipedia page, http://en.wikipedia.org/wiki/SNPedia * Download URL: http://www.SNPedia.com/index.php/Bulk * Web Service URL: http://bots.SNPedia.com/api.php
Proper citation: SNPedia (RRID:SCR_006125) Copy
http://research.nhgri.nih.gov/CGD/
Manually curated database of all conditions with known genetic causes, focusing on medically significant genetic data with available interventions. Includes gene symbol, conditions, allelic conditions, inheritance, age in which interventions are indicated, clinical categorization, and general description of interventions/rationale. Contents are intended to describe types of interventions that might be considered. Includes only single gene alterations and does not include genetic associations or susceptibility factors related to more complex diseases.
Proper citation: Clinical Genomic Database (RRID:SCR_006427) Copy
https://github.com/MicrosoftGenomics/FaST-LMM
FaST-LMM (Factored Spectrally Transformed Linear Mixed Models) is a set of tools for efficiently performing genome-wide association studies (GWAS), prediction, and heritability estimation on large data sets.
Proper citation: FaST LMM (RRID:SCR_015506) Copy
https://github.com/gatech-genemark/ProtHint
Software pipeline for predicting and scoring hints (in form of introns, start and stop codons) in genome of interest by mapping and spliced aligning predicted genes to database of reference protein sequences.
Proper citation: ProtHint (RRID:SCR_021167) Copy
http://ccr.coriell.org/Sections/Collections/HuREF/?SsId=78
The Human Reference Genetic Material Repository makes available DNA from a single individual, J. Craig Venter, whose genome has been sequenced and assembled. The DNA samples are prepared from a lymphoblastoid cell line established at Coriell Cell Repositories from a sample of peripheral blood. The DNA samples are available in 50 microgram aliquots. The lymphoblastoid cell line is not available for distribution. The human DNA sample provided is that of J. Craig Venter whose DNA from white blood cells and sperm was sequenced using Sanger chemistry (ABI Capillary Electrophoresis Platforms 3700 and 3730xl), assembled using the Celera Assembler and was published in PLoS Biology . J. Craig Venter, born on 14 October 1946, is a Caucasian male of self-reported European-American ancestry. The data available on this sample, whose genome assembly is referred to as HuRef, includes: * Whole Genome Shotgun Sequencing data * Sequence trace set deposited by JCVI in the NCBI trace archive * Human Genome Browser displaying sequence assembly, DNA variants and gene annotations Additional data sets from this study include: * Full set of Sanger reads used for genome assembly * SNP and insertion/deletion variant on the human genome sequence coordinates (NCBI version 36) * Affymetrix 500K GeneChip data * Illumina HumanHap650Y Genotyping BeadChip data Given the amount of data publicly available the genomic content of this sample, HuRef will be useful as a reference for many genetic studies.
Proper citation: Human Reference Genetic Material Repository (RRID:SCR_004693) Copy
Part of zebrafish genome project. ZGC project to produce cDNA libraries, clones and sequences to provide complete set of full-length (open reading frame) sequences and cDNA clones of expressed genes for zebrafish. All ZGC sequences are deposited in GenBank and clones can be purchased from distributors of IMAGE consortium. With conclusion of ZGC project in September 2008, GenBank records of ZGC sequences will be frozen, without further updates. Since definition of what constitutes full-length coding region for some of genes and transcripts for which we have ZGC clones will likely change in future, users planning to order ZGC clones will need to monitor for these changes. Users can make use of genome browsers and gene-specific databases, such as UCSC Genome browser, NCBI's Map Viewer, and Entrez Gene, to view relevant regions of genome (browsers) or gene-related information (Entrez Gene).
Proper citation: Zebrafish Gene Collection (RRID:SCR_007054) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the SPARC SAWG Resources search. From here you can search through a compilation of resources used by SPARC SAWG and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that SPARC SAWG has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on SPARC SAWG then you can log in from here to get additional features in SPARC SAWG such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into SPARC SAWG you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within SPARC SAWG that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.