Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
Open, web-based platform providing bioinformatics tools and services for data intensive genomic research. Platform may be used as a service or installed locally to perform, reproduce, and share complete analyses. Galaxy automatically tracks and manages data provenance and provides support for capturing the context and intent of computational methods. Galaxy Community has created Galaxy instances in many different forms and for many different applications including Galaxy servers, cloud services that support Galaxy instances, and virtual machines and containers that can be easily deployed for your own server.The Galaxy team is a part of BX at Penn State, and the Biology and Mathematics and Computer Science departments at Emory University.Training Infrastructure as a Service (TIaaS) is a service offered by some UseGalaxy servers to specifically support training use cases.
Proper citation: Galaxy (RRID:SCR_006281) Copy
http://lincs.hms.harvard.edu/db/
Database that contains all publicly available HMS LINCS datasets and information for each dataset about experimental reagents and experimental and data analysis protocols. Experimental reagents include small molecule perturbagens, cells, antibodies, and proteins.
Proper citation: HMS LINCS Database (RRID:SCR_006454) Copy
Set of measures intended for use in large-scale genomic studies. Facilitate replication and validation across studies. Includes links to standards and resources in effort to facilitate data harmonization to legacy data. Measurement protocols that address wide range of research domains. Information about each protocol to ensure consistent data collection.Collections of protocols that add depth to Toolkit in specific areas.Tools to help investigators implement measurement protocols.
Proper citation: Phenotypes and eXposures Toolkit (RRID:SCR_006532) Copy
http://www.informatics.jax.org/searches/AMA_form.shtml
Ontology that organizes anatomical structures for the adult mouse (Theiler stage 28) spatially and functionally, using ''is a'' and ''part of'' relationships. The ontology is used to describe expression data for the adult mouse and phenotype data pertinent to anatomy in standardized ways. The browser can be used to view anatomical terms and their relationships in a hierarchical display.
Proper citation: Adult Mouse Anatomy Ontology (RRID:SCR_006568) Copy
http://dgidb.genome.wustl.edu/
A database of drug-gene relationships that provides drug-gene interactions and potential druggability data given list of genes. There are about 15 data sources that are being aggregated by DGIdb, with update date and these data sources are listed on this page: http://dgidb.genome.wustl.edu/sources, THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: DGIdb (RRID:SCR_006608) Copy
Encyclopedia of DNA elements consisting of list of functional elements in human genome, including elements that act at protein and RNA levels, and regulatory elements that control cells and circumstances in which gene is active. Enables scientific and medical communities to interpret role of human genome in biology and disease. Provides identification of common cell types to facilitate integrative analysis and new experimental technologies based on high-throughput sequencing. Genome Browser containing ENCODE and Epigenomics Roadmap data. Data are available for entire human genome.
Proper citation: ENCODE (RRID:SCR_006793) Copy
International collaboration producing an extensive public catalog of human genetic variation, including SNPs and structural variants, and their haplotype contexts, in an effort to provide a foundation for investigating the relationship between genotype and phenotype. The genomes of about 2500 unidentified people from about 25 populations around the world were sequenced using next-generation sequencing technologies. Redundant sequencing on various platforms and by different groups of scientists of the same samples can be compared. The results of the study are freely and publicly accessible to researchers worldwide. The consortium identified the following populations whose DNA will be sequenced: Yoruba in Ibadan, Nigeria; Japanese in Tokyo; Chinese in Beijing; Utah residents with ancestry from northern and western Europe; Luhya in Webuye, Kenya; Maasai in Kinyawa, Kenya; Toscani in Italy; Gujarati Indians in Houston; Chinese in metropolitan Denver; people of Mexican ancestry in Los Angeles; and people of African ancestry in the southwestern United States. The goal Project is to find most genetic variants that have frequencies of at least 1% in the populations studied. Sequencing is still too expensive to deeply sequence the many samples being studied for this project. However, any particular region of the genome generally contains a limited number of haplotypes. Data can be combined across many samples to allow efficient detection of most of the variants in a region. The Project currently plans to sequence each sample to about 4X coverage; at this depth sequencing cannot provide the complete genotype of each sample, but should allow the detection of most variants with frequencies as low as 1%. Combining the data from 2500 samples should allow highly accurate estimation (imputation) of the variants and genotypes for each sample that were not seen directly by the light sequencing. All samples from the 1000 genomes are available as lymphoblastoid cell lines (LCLs) and LCL derived DNA from the Coriell Cell Repository as part of the NHGRI Catalog. The sequence and alignment data generated by the 1000genomes project is made available as quickly as possible via their mirrored ftp sites. ftp://ftp.1000genomes.ebi.ac.uk ftp://ftp-trace.ncbi.nlm.nih.gov/1000genomes
Proper citation: 1000 Genomes: A Deep Catalog of Human Genetic Variation (RRID:SCR_006828) Copy
http://biositemaps.ncbcs.org/rds/search.html
Resource Discovery System is a web-accessible and searchable inventory of biomedical research resources. Powered by the Resource Discovery System (RDS) that includes a standards-based informatics infrastructure * Biositemaps Information Model * Biomedical Resource Ontology Extensions * Web Services distributed web-accessible inventory framework * Biositemap Resource Editor * Resource Discovery System Source code and project documentation to be made available on an open-source basis. Contributing institutions: University of Pittsburgh, University of Michigan, Stanford University, Oregon Health & Science University, University of Texas Houston. Duke University, Emory University, University of California Davis, University of California San Diego, National Institutes of Health, Inventory Resources Working Group Members
Proper citation: Resource Discovery System (RRID:SCR_005554) Copy
Collects mammalian cis- and trans-regulatory elements together with experimental evidence. Regulatory elements were mapped on to assembled genomes. Resource for gene regulation and function studies. Users can retrieve primers, search TF target genes, retrieve TF motifs, search Gene Regulatory Networks and orthologs, and make use of sequence analysis tools. Uses databases such as Genbank, EPD and DBTSS, and employ promoter finding program FirstEF combined with mRNA/EST information and cross-species comparisons. Manually curated.
Proper citation: Transcriptional Regulatory Element Database (RRID:SCR_005661) Copy
http://www.jcvi.org/charprotdb/index.cgi/home
The Characterized Protein Database, CharProtDB, is designed and being developed as a resource of expertly curated, experimentally characterized proteins described in published literature. For each protein record in CharProtDB, storage of several data types is supported. It includes functional annotation (several instances of protein names and gene symbols) taxonomic classification, literature links, specific Gene Ontology (GO) terms and GO evidence codes, EC (Enzyme Commisssion) and TC (Transport Classification) numbers and protein sequence. Additionally, each protein record is associated with cross links to all public accessions in major protein databases as ��synonymous accessions��. Each of the above data types can be linked to as many literature references as possible. Every CharProtDB entry requires minimum data types to be furnished. They are protein name, GO terms and supporting reference(s) associated to GO evidence codes. Annotating using the GO system is of importance for several reasons; the GO system captures defined concepts (the GO terms) with unique ids, which can be attached to specific genes and the three controlled vocabularies of the GO allow for the capture of much more annotation information than is traditionally captured in protein common names, including, for example, not just the function of the protein, but its location as well. GO evidence codes implemented in CharProtDB directly correlate with the GO consortium definitions of experimental codes. CharProtDB tools link characterization data from multiple input streams through synonymous accessions or direct sequence identity. CharProtDB can represent multiple characterizations of the same protein, with proper attribution and links to database sources. Users can use a variety of search terms including protein name, gene symbol, EC number, organism name, accessions or any text to search the database. Following the search, a display page lists all the proteins that match the search term. Click on the protein name to view more detailed annotated information for each protein. Additionally, each protein record can be annotated.
Proper citation: CharProtDB: Characterized Protein Database (RRID:SCR_005872) Copy
http://www.neuroepigenomics.org/methylomedb/
A database containing genome-wide brain DNA methylation profiles for human and mouse brains. The DNA methylation profiles were generated by Methylation Mapping Analysis by Paired-end Sequencing (Methyl-MAPS) method and analyzed by Methyl-Analyzer software package. The methylation profiles cover over 80% CpG dinucleotides in human and mouse brains in single-CpG resolution. The integrated genome browser (modified from UCSC Genome Browser allows users to browse DNA methylation profiles in specific genomic loci, to search specific methylation patterns, and to compare methylation patterns between individual samples. Two species were included in the Brain Methylome Database: human and mouse. Human postmortem brain samples were obtained from three distinct cortical regions, i.e., dorsal lateral prefrontal cortex (dlPFC), ventral prefrontal cortex (vPFC), and auditory cortex (AC). Human samples were selected from our postmortem brain collection with extensive neuropathological and psychopathological data, as well as brain toxicology reports. The Department of Psychiatry of Columbia University and the New York State Psychiatric Institute have assembled this brain collection, where a validated psychological autopsy method is used to generate Axis I and II DSM IV diagnoses and data are obtained on developmental history, history of psychiatric illness and treatment, and family history for each subject. The mouse sample (strain 129S6/SvEv) DNA was collected from the entire left cerebral hemisphere. The three human brain regions were selected because they have been implicated in the neuropathology of depression and schizophrenia. Within each cortical region, both disease and non-psychiatric samples have been profiled (matching subjects by age and sex in each group). Such careful matching of subjects allows one to perform a wide range of queries with the ability to characterize methylation features in non-psychiatric controls, as well as detect differentially methylated domains or features between disease and non-psychiatric samples. A total of 14 non-psychiatric, 9 schizophrenic, and 6 depression methylation profiles are included in the database.
Proper citation: MethylomeDB (RRID:SCR_005583) Copy
http://www.broadinstitute.org/mammals/haploreg/haploreg.php
HaploReg is a tool for exploring annotations of the noncoding genome at variants on haplotype blocks, such as candidate regulatory SNPs at disease-associated loci. Using linkage disequilibrium (LD) information from the 1000 Genomes Project, linked SNPs and small indels can be visualized along with their predicted chromatin state in nine cell types, conservation across mammals, and their effect on regulatory motifs. HaploReg is designed for researchers developing mechanistic hypotheses of the impact of non-coding variants on clinical phenotypes and normal variation.
Proper citation: HaploReg (RRID:SCR_006796) Copy
http://hanalyzer.sourceforge.net/
An open-source data integration system designed to assist biologists in explaining the results observed in genome-scale experiments as well as generating new hypotheses. It combines information extraction techniques, semantic data integration, and reasoning and facilitates network visualization. The Hanalyzer source code and binaries are available for download.
Proper citation: Hanalyzer (RRID:SCR_000923) Copy
Database of Drosophila transcription factor DNA binding specificity using the bacterial one-hybrid method, DNase I or SELEX methods. The database provides community access to recognition motifs and position weight matrices for transcription factors (TFs), including many unpublished motifs. Search tools and flat file downloads are provided to retrieve binding site information (as sequences, matrices and sequence logos) for individual TFs, groups of TFs or for all TFs with characterized binding specificities. Linked analysis tools allow users to identify motifs within the database that share similarity to a query matrix or to view the distribution of occurrences of an individual motif throughout the Drosophila genome. This database and its associated tools provide computational and experimental biologists with resources to predict interactions between Drosophila TFs and target cis-regulatory sequences.
Proper citation: FlyFactorSurvey (RRID:SCR_002113) Copy
Model organism database that serves as central repository and web-based resource for zebrafish genetic, genomic, phenotypic and developmental data. Data represented are derived from three primary sources: curation of zebrafish publications, individual research laboratories and collaborations with bioinformatics organizations. Data formats include text, images and graphical representations.Serves as primary community database resource for laboratory use of zebrafish. Developed and supports integrated zebrafish genetic, genomic, developmental and physiological information and link this information extensively to corresponding data in other model organism and human databases.
Proper citation: Zebrafish Information Network (ZFIN) (RRID:SCR_002560) Copy
http://sonorus.princeton.edu/hefalmp/
HEFalMp (Human Experimental/FunctionAL MaPper) is a tool developed by Curtis Huttenhower in Olga Troyanskaya's lab at Princeton University. It was created to allow interactive exploration of functional maps. Functional mapping analyzes portions of these networks related to user-specified groups of genes and biological processes and displays the results as probabilities (for individual genes), functional association p-values (for groups of genes), or graphically (as an interaction network). HEFalMp contains information from roughly 15,000 microarray conditions, over 15,000 publications on genetic and physical protein interactions, and several types of DNA and protein sequence analyses and allows the exploration of over 200 H. sapiens process-specific functional relationship networks, including a global, process-independent network capturing the most general functional relationships. Looking to download functional maps? Keep an eye on the bottom of each page of results: every functional map of any kind is generated with a Download link at the bottom right. Most functional maps are provided as tab-delimited text to simplify downstream processing; graphical interaction networks are provided as Support Vector Graphics files, which can be viewed using the Adobe Viewer, any recent version of Firefox, or the excellent open source Inkscape tool.
Proper citation: Human Experimental/FunctionAL MaPper: Providing Functional Maps of the Human Genome (RRID:SCR_003506) Copy
https://www.hsph.harvard.edu/alkes-price/software/
Software application that uses genotyping data from SNP arrays for accurately inferring chromosomal segments of distinct continental ancestry in admixed populations, using dense genetic data. (entry from Genetic Analysis Software)
Proper citation: Hapmix (RRID:SCR_004203) Copy
An experiment in web-database access to large multi-dimensional data sets using a standardized experimental platform to determine if the larger scientific community can be given simple, intuitive, and user-friendly web-based access to large microarray data sets. All data in PEPR is also available via NCBI GEO. The structure and goals of PEPR differ from other mRNA expression profiling databases in a number of important ways. * The experimental platform in PEPR is standardized, and is an Affymetrix - only database. All microarrays available in the PEPR web database should ascribe to quality control and standard operating procedures. A recent publication has described the QC/SOP criteria utilized in PEPR profiles ( The Tumor Analysis Best Practices Working Group 2004 ). * PEPR permits gene-based queries of large Affymetrix array data sets without any specialized software. For example, a number of large time series projects are available within PEPR, containing 40-60 microarrays, yet these can be simply queried via a dynamic web interface with no prior knowledge of microarray data analysis. * Projects in PEPR originate from scientists world-wide, but all data has been generated by the Research Center for Genetic Medicine, Children''''s National Medical Center, Washington DC. Future developments of PEPR will allow remote entry of Affymetrix data ascribing to the same QC/SOP protocols. They have previously described an initial implementation of PEPR, and a dynamic web-queried time series graphical interface ( Chen et al. 2004 ). A publication showing the utility of PEPR for pharmacodynamic data has recently been published ( Almon et al. 2003 ).
Proper citation: Public Expression Profiling Resource (RRID:SCR_007274) Copy
http://www.oreganno.org/oregano/
Open source, open access database and literature curation system for community based annotation of experimentally identified DNA regulatory regions, transcription factor binding sites and regulatory variants. Automatically cross referenced against PubMED, Entrez Gene, EnsEMBL, dbSNP, eVOC: Cell type ontology, and Taxonomy database. Community driven resource for curated regulatory annotation.
Proper citation: Open Regulatory Annotation Database (RRID:SCR_007835) Copy
http://www.uniprot.org/help/uniref
Databases which provide clustered sets of sequences from UniProt Knowledgebase and selected UniParc records, in order to obtain complete coverage of sequence space at several resolutions while hiding redundant sequences from view. The UniRef100 database combines identical sequences and sub-fragments with 11 or more residues (from any organism) into a single UniRef entry. The sequence of a representative protein, the accession numbers of all the merged entries, and links to the corresponding UniProtKB and UniParc records are all displayed in the entry. UniRef90 and UniRef50 are built by clustering UniRef100 sequences with 11 or more residues such that each cluster is composed of sequences that have at least 90% (UniRef90) or 50% (UniRef50) sequence identity to the longest sequence (UniRef seed sequence). All the sequences in each cluster are ranked to facilitate the selection of a representative sequence for the cluster.
Proper citation: UniRef (RRID:SCR_010646) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the kravitz2 Resources search. From here you can search through a compilation of resources used by kravitz2 and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that kravitz2 has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on kravitz2 then you can log in from here to get additional features in kravitz2 such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into kravitz2 you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within kravitz2 that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.