Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://sisyphus.mrc-cpe.cam.ac.uk
THIS RESOURCE IS NO LONGER IN SERVICE, documented on July 15, 2013. A collection of manually curated protein structural alignments and their interrelationships. Each multiple alignment within the SISYPHUS database consists of structurally similar regions common to a group of proteins. These regions range from oligomeric biological units, or individual domains to fragments of different size representing either internal structural repeats or motifs common to structurally distinct proteins. The SISYPHUS multiple alignments are displayed with SPICE, a browser that provides an integrated view of protein sequences, structures and their annotations.
Proper citation: SISYPHUS (RRID:SCR_007930) Copy
http://silkworm.genomics.org.cn/
THIS RESOURCE IS NO LONGER IN SERVICE, documented May 10, 2017. A pilot effort that has developed a centralized, web-based biospecimen locator that presents biospecimens collected and stored at participating Arizona hospitals and biospecimen banks, which are available for acquisition and use by researchers. Researchers may use this site to browse, search and request biospecimens to use in qualified studies. The development of the ABL was guided by the Arizona Biospecimen Consortium (ABC), a consortium of hospitals and medical centers in the Phoenix area, and is now being piloted by this Consortium under the direction of ABRC. You may browse by type (cells, fluid, molecular, tissue) or disease. Common data elements decided by the ABC Standards Committee, based on data elements on the National Cancer Institute''s (NCI''s) Common Biorepository Model (CBM), are displayed. These describe the minimum set of data elements that the NCI determined were most important for a researcher to see about a biospecimen. The ABL currently does not display information on whether or not clinical data is available to accompany the biospecimens. However, a requester has the ability to solicit clinical data in the request. Once a request is approved, the biospecimen provider will contact the requester to discuss the request (and the requester''s questions) before finalizing the invoice and shipment. The ABL is available to the public to browse. In order to request biospecimens from the ABL, the researcher will be required to submit the requested required information. Upon submission of the information, shipment of the requested biospecimen(s) will be dependent on the scientific and institutional review approval. Account required. Registration is open to everyone.. Documented on August 20,2019.A database of integrated genome resources for the silkworm, Bombyx mori. This database provides access to not only genomic data including functional annotation of genes, gene products and chromosomal mapping, but also extensive biological information such as microarray expression data, ESTs and corresponding references. SilkDB will be useful for the silkworm research community as well as comparative genomics. Recently, an international collaboration has been launched to assemble a complete silkworm genome sequence, which is based on the 6� and 3� draft genome sequences created by Chinese group and Japanese group in 2004 (Mita et al., 2004; Xia et al., 2004), respectively. The genome assembly quality has been greatly improved. Base on a high density SNP genetic map, over 80% of genome sequence could be mapped on 28 chromosomes of the silkworm. The first version of SilkDB was released in 2004. Since that time, the silkworm has become a focus in insect research community and the study of silkworm has been greatly accelerated. Now, we are happy to announce the release of a new version of SilkDB, which updated all of the data, added new information of genome sequence and genes, and provides new tools to facilitate use of the genome database.
Proper citation: SilkDB (RRID:SCR_007926) Copy
Database that provides access to mRNA sequences and associated regulatory elements that were processed from Genbank. These mRNA sequences include complete genomes, which are divided into 5-prime UTRs, 3-prime UTRs, initiation sequences, termination regions and full CDS sequences. This data can be searched for a range of properties including specific mRNA sequences, mRNA motifs, codon usage, RSCU values, information content, etc.
Proper citation: Transterm (RRID:SCR_008244) Copy
http://pbil.univ-lyon1.fr/databases/homolens.php
Database of homologous genes from Ensembl organisms, structured under ACNUC sequence database management system. It allows to select sets of homologous genes among species, and to visualize multiple alignments and phylogenetic trees. It is possible to search for orthologous genes in a wide range of taxons. HOMOLENS is particularly useful for comparative sequence analysis, phylogeny and molecular evolution studies. More generally, HOMOLENS gives an overall view of what is known about a peculiar gene family. Note that HOMOLENS is split into two databases on this server: HOMOLENS contains the protein sequences while HOMOLENSDNA contains the nucleotide sequences. Protein sequences of HOMOLENS have been generated by translating the CDS of HOMOLENSDNA and using associated cross-references to generate the annotations.
Proper citation: Homologous Sequences in Ensembl Animal Genomes (RRID:SCR_008356) Copy
http://www.bioinformatics2.wsu.edu/cgi-bin/Athena/cgi/home.pl
Athena is a web-based application that warehouses disparate datatypes related to the control of gene expression. Athena provides several features to enable exploration of the regulatory mechanisms of Arabidopsis gene control. The first main tool we provide is visualization of promoter domains of selected genes. Database crossreference for these transcription factors is provided as well as a statistical test for enrichment of binding activity within the set of selected promoters. The data mining tools in Athena allow for selection of sets of genes based on two different factors. -Genes can be select by specifying a set of binding factors whose putative sites must be present within all of those genes'' promoter regions. -Alternatively, genes can be selected using Gene Ontology annotations. Both GO (Gene Ontology) Slim terms and Gene Ontology terms are available. One can select a set of genes by either choosing a union of the genes annotated by a selected set of Slim terms or Gene Ontology terms. The selected gene''s putative binding factors are listed, including enrichment data. Furthermore, enriched presence of Gene Ontology terms is given. The analysis suite provides both enhanced data mining tools for selecting genes as well as several data displays., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Athena (RRID:SCR_008110) Copy
http://andromeda.gsf.de/litminer
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 23, 2016. The LitMiner software is a literature data-mining tool that facilitates the identification of major gene regulation key players related to a user-defined field of interest in PubMed abstracts. The prediction of gene-regulatory relationships is based on co-occurrence analysis of key terms within the abstracts. LitMiner predicts relationships between key terms from the biomedical domain in four categories (genes, chemical compounds, diseases and tissues). The usefulness of the LitMiner system has been demonstrated recently in a study that reconstructed disease-related regulatory networks by promoter modeling that was initiated by a LitMiner generated primary gene list. To overcome the limitations and to verify and improve the data, we developed WikiGene, a Wiki-based curation tool that allows revision of the data by expert users over the Internet. It is based on the annotation of key terms in article abstracts followed by statistical co-citation analysis of annotated key terms in order to predict relationships. Key terms belonging to four different categories are used for the annotation process: -Genes: Names of genes and gene products. Gene name recognition is based on Ensembl . Synonyms and aliases are resolved. -Chemical Compounds: Names of chemical compounds and their respective aliases. -Diseases and Phenotypes: Names of diseases and phenotypes -Tissues and Organs: Names of tissues and organs LitMiner uses a database of disease and phenotype terms for literature annotation. Currently, there are 2225 diseases or phenotypes, 801 tissues and organs, and 10477 compounds in the database.
Proper citation: LitMiner (RRID:SCR_008200) Copy
http://www.cmbi.ru.nl/GeneSeeker/
The GeneSeeker allows you to search across different databases simultaneously, given a known human genetic location and expression/phenotypic pattern. The GeneSeeker returns any found gene names which are located on the specified location and expressed in the specified tissue. To search for more expression location in one search, just enter them in the textbox for the expression location and separate them with logical operators (and, or, not). You can specify as many tissues as you want, the program starts 20 queries simultaneously, and then waits for a query to finish before starting another query, to keep server loads to a minimum. You can also search only for expression, just leave the cytogenetic location fields blank, and do the query. If you only want to look for one cytogenetic location, only fill in the first location field, and the GeneSeeker will search with only this one. Housekeeping genes , found in Swissprot can be excluded, or genes that are to be excluded can be specified. Human chromosome localizations are translated with an oxford-grid to mouse chromosome localizations, and then submitted to the Mgd. Sponsors: GeneSeeker is a service provided by the Centre for Molecular and Biomolecular Informatics (CMBI).
Proper citation: GeneSeeker (RRID:SCR_008347) Copy
http://www.ebi.ac.uk/thornton-srv/databases/WSsas/
SAS is a tool for applying structural information to a given protein sequence. It uses FASTA to scan a given protein sequence against all the proteins of known 3D structure in the Protein Data Bank and provides functional residue annotation based on data from the Catalytic Site Atlas and PDBsum. The web service is aimed to facilitate the use of the SAS tool when having a huge number of queries. Currently, the web service provides annotation for binding sites (to ligand, metal or nucleic acid), catalytic residues and amino acids related to protein-protein interactions.
Proper citation: WSsas - Web Service for the SAS tool (RRID:SCR_007051) Copy
http://autismkb.cbi.pku.edu.cn/
Genetic factors contribute significantly to ASD. AutismKB is an evidence-based knowledgebase of Autism spectrum disorder (ASD) genetics. The current version contains 2193 genes (99 syndromic autism related genes and 2135 non-syndromic autism related genes), 4617 Copy Number Variations (CNVs) and 158 linkage regions associated with ASD by one or more of the following six experimental methods: # Genome-Wide Association Studies (GWAS); # Genome-wide CNV studies; # Linkage analysis; # Low-scale genetic association studies; # Expression profiling; # Other low-scale gene studies. Based on a scoring and ranking system, 99 syndromic autism related genes and 383 non-syndromic autism related genes (434 genes in total) were designated as having high confidence. Autism spectrum disorder (ASD) is a heterogeneous neurodevelopmental disorder with a prevalence of 1.0-2.6%. The three core symptoms of ASD are: # impairments in reciprocal social interaction; # communication impairments; # presence of restricted, repetitive and stereotyped patterns of behavior, interests and activities.
Proper citation: AutismKB (RRID:SCR_006937) Copy
https://github.com/jstjohn/SimSeq
An illumina paired-end and mate-pair short read simulator. This project attempts to model as many of the quirks that exist in Illumina data as possible. Some of these quirks include the potential for chimeric reads, and non-biotinylated fragment pull down in mate-pair libraries .
Proper citation: SimSeq (RRID:SCR_006947) Copy
http://weizhong-lab.ucsd.edu/cd-hit/
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute)., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: CD-HIT (RRID:SCR_007105) Copy
This database provides a platform to query and compare gene expression data during the development of the major model animals (zebrafish, drosophila, medaka, mouse). The name 4DXpress stands for expression database in 4D. The 4D (four dimensions) of 4DXpress can be interpreted either as: 3 spatial dimensions plus time, or as 1. species 2. gene 3. developmental stage 4. anatomical structure. The major focus of this database lies in cross species comparison. The high resolution expression data was acquired through whole mount in situ hybridsation-, antibody- or transgenic experiments. Data was integrated from several species specific expression pattern databases, such as ZFIN, BDGP, GXD, MEPD as well as directly submitted by researchers of the participating groups at EMBL. The 4DXpress database is a project within the Centre for Computational Biology at EMBL. It is developed by Yannick Haudry, Thorsten Henrich and Ivica Letunic and coordinated by Thorsten Henrich. Hugo Berube is developing the 4D ArrayExpress Data Warehouse at EBI for integrating in situ data with microarray data.
Proper citation: Expression Database in 4D (RRID:SCR_007066) Copy
Database containing the DNA sequence and annotation of the entire human chromosome 7, encompassing nearly 158 million nucleotides of DNA and 1917 gene structures, are presented; the most up to date collation of sequence, gene, and other annotations from all databases (eg. Celera published, NCBI, Ensembl, RIKEN, UCSC) as well as unpublished data. To generate a higher order description, additional structural features such as imprinted genes, fragile sites, and segmental duplications were integrated at the level of the DNA sequence with medical genetic data, including 440 chromosome rearrangement breakpoints associated with disease. The objective of this project is to generate a comprehensive description of human chromosome 7 to facilitate biological discovery, disease gene research and medical genetic applications. There are over 360 disease-associated genes or loci on chromosome 7. A major challenge ahead will be to represent chromosome alterations, variants, and polymorphisms and their related phenotypes (or lack thereof), in an accessible way. In addition to being a primary data source, this site serves as a weighing station for testing community ideas and information to produce highly curated data to be submitted to other databases such as NCBI, Ensembl, and UCSC. Therefore, any useful data submitted will be curated and shown in this database. All Chromosome 7 genomic clones (cosmids, BACs, YACs) listed in GBrowser and in other data tables are freely distributed.
Proper citation: Chromosome 7 Annotation Project (RRID:SCR_007134) Copy
Software platform, general technologies and theoretical supports for computational biology with the grand aim to make precise whole cell simulation at the molecular level possible.Technologies include formalisms and techniques, including technologies to predict, obtain or estimate parameters such as reaction rates and concentrations of molecules in the cell. The E-Cell System is a software platform for modeling, simulation and analysis of complex, heterogeneous and multi-scale system like the cell. The E-Cell Project is open to anyone who shares the view with u that development of cell simulation technology, and, even if such ultimate goal might not be within ten years of reach yet, solving various conceptual, computational and experimental problems that will continue to arise in the course of pursuing it, may have a multitude of eminent scientific, medical and engineering impacts on our society.
Proper citation: Electronic Cell Project (RRID:SCR_007381) Copy
Resource for experimentally validated human and mouse noncoding fragments with gene enhancer activity as assessed in transgenic mice. Most of these noncoding elements were selected for testing based on their extreme conservation in other vertebrates or epigenomic evidence (ChIP-Seq) of putative enhancer marks. Central public database of experimentally validated human and mouse noncoding fragments with gene enhancer activity as assessed in transgenic mice. Users can retrieve elements near single genes of interest, search for enhancers that target reporter gene expression to particular tissue, or download entire collections of enhancers with defined tissue specificity or conservation depth.
Proper citation: VISTA Enhancer Browser (RRID:SCR_007973) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on August 26,2019. In October 2016, T1DBase has merged with its sister site ImmunoBase (https://immunobase.org). Documented on March 2020, ImmunoBase ownership has been transferred to Open Targets (https://www.opentargets.org). Results for all studies can be explored using Open Targets Genetics (https://genetics.opentargets.org). Database focused on genetics and genomics of type 1 diabetes susceptibility providing a curated and integrated set of datasets and tools, across multiple species, to support and promote research in this area. The current data scope includes annotated genomic sequences for suspected T1D susceptibility regions; genetic data; microarray data; and global datasets, generally from the literature, that are useful for genetics and systems biology studies. The site also includes software tools for analyzing the data.
Proper citation: T1DBase (RRID:SCR_007959) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 23,2023.Software package for comparison and analysis of microbial communities, primarily based on high-throughput amplicon sequencing data, but also supporting analysis of other types of data. QIMME analyzes and transforms raw sequencing data generated on Illumina or other platforms to publication quality graphics and statistics.
Proper citation: QIIME (RRID:SCR_008249) Copy
https://www.ncbi.nlm.nih.gov/genbank/dbest/
Database as a division of GenBank that contains sequence data and other information on single-pass cDNA sequences, or Expressed Sequence Tags, from a number of organisms.
Proper citation: dbEST (RRID:SCR_008132) Copy
http://www.baderlab.org/Software/ActiveDriver
A statistical method for interpreting variations in protein sequence (e.g. coding SNPs in the population, SNVs in cancer genomes) in the context of protein post-translational signaling modifications.
Proper citation: ActiveDriver (RRID:SCR_008104) Copy
http://tree.bio.ed.ac.uk/software/figtree
A graphical viewer of phylogenetic trees and a program for producing publication-ready figures. It is designed to display summarized and annotated trees produced by BEAST.
Proper citation: FigTree (RRID:SCR_008515) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the SPARC SAWG Resources search. From here you can search through a compilation of resources used by SPARC SAWG and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that SPARC SAWG has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on SPARC SAWG then you can log in from here to get additional features in SPARC SAWG such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into SPARC SAWG you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within SPARC SAWG that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.