Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://metaphyler.cbcb.umd.edu/
A taxonomic classifier for metagenomic shotgun reads, which uses phylogenetic marker genes as a taxonomic reference. The classifier, based on BLAST, uses different thresholds (automatically learned from the reference database) for each combination of taxonomic rank, reference gene, and sequence length. The reference database includes marker genes from all complete genomes, several draft genomes and the NCBI nr protein database.
Proper citation: MetaPhyler (RRID:SCR_004848) Copy
http://blast.ncbi.nlm.nih.gov/Blast.cgi
Web search tool to find regions of similarity between biological sequences. Program compares nucleotide or protein sequences to sequence databases and calculates statistical significance. Used for identifying homologous sequences.
Proper citation: NCBI BLAST (RRID:SCR_004870) Copy
http://www.well.ox.ac.uk/~kgaulton/chaos.shtml
A Perl-based system for annotation of variants identified in high-throughput sequencing experiments. Functionality includes annotation of variants with information relating to population genetics, known transcripts, positional records, and sequence motif-based prediction. In addition, annotated variants can be summarized and extracted to facilitate downstream analysis. There is also basic support for gene-based biological annotation, and eventually will include tools for variant and genotype analysis and visualization.
Proper citation: CHAoS (RRID:SCR_005174) Copy
http://cbrc.kaust.edu.sa/readscan/
A highly scalable parallel software program to identify non-host sequences (of potential pathogen origin) and estimate their genome relative abundance in high-throughput sequence datasets.
Proper citation: READSCAN (RRID:SCR_005204) Copy
A web server designed to rapidly and accurately identify, annotate and graphically display prophage sequences within bacterial genomes or plasmids. It accepts either raw DNA sequence data or partially annotated GenBank formatted data and rapidly performs a number of database comparisons as well as phage cornerstone feature identification steps to locate, annotate and display prophage sequences and prophage features. Relative to other prophage identification tools, PHAST is up to 40 times faster and up to 15% more sensitive. It is also able to process and annotate both raw DNA sequence data and Genbank files, provide richly annotated tables on prophage features and prophage quality and distinguish between intact and incomplete prophage. PHAST also generates downloadable, high quality, interactive graphics that display all identified prophage components in both circular and linear genomic views. Databases available for download include Virus DB, Prophage and virus DB, Bacteria DB, and PHAST result DB. Pre-calculated genomes for viewing are also available.
Proper citation: PHAge Search Tool (RRID:SCR_005184) Copy
Portal supporting the North East Bioinformatics Collaborative''s project to sequence the genome of the Little Skate. Provided is a clearinghouse for Little Skate Genome Project and other publicly available Skate and Ray (Batoidea) genome data, and tools for data visualization and analysis. Little Skate Genome Project The little skate (Leucoraja erinacea) is a chondrichthyan (cartilaginous) fish native to the east coast of North America. Elasmobranchs (Skates, Rays, and Sharks) exhibit many fundamental vertebrate characteristics, including a neural crest, jaws and teeth, an adaptive immune system, and a pressurized circulatory system. These characteristics have been exploited to promote understanding about human physiology, immunology, stem cell biology, toxicology, neurobiology and regeneration. The development of standardized experimental protocols in elasmobranchs such as L. erinacea and the spiny dogfish shark (Squalus acanthias) has further positioned these organisms as important biomedical and developmental models. Despite this distinction, the only reported chondrichthyan genome is the low coverage (1.4x) draft genome of the elephant shark (Callorhinchus milii). To close the evolutionary gaps in available elasmobranch genome sequence data, and generate critical genomic resources for future biomedical study, the genome of L. erinacea is being sequenced by the North East Bioinformatics Collaborative (NEBC). As close evolutionary relatives, the little skate sequence will facilitate studies that employ dogfish shark and other elasmobranchs as model organisms. Skate tools include the SkateBLAST and the Skate Genome Browsers: Little Skate Mitochondrion, Thorny Skate Mitochondrion, and Ocellate Spot Skate Mitochondrion.
Proper citation: SkateBase (RRID:SCR_005302) Copy
http://www.glycosciences.de/modeling/sweet2/
Program that rapidly converts the primary sequence of a complex carbohydrate, as defined by standard nomenclature, directly into a reliable 3D molecular model by linking together preconstructed 3D molecular templates of monosaccharides in the manner specified by the sequence and then optimizing the 3D structure using the MM3 force field. The user interaction is supported by an input spreadsheet consisting of a grid of sugar symbol and connection type cells. Several ways to visualize and to output the generated structures and related information are implemented.
Proper citation: SWEET-DB (RRID:SCR_005324) Copy
http://www.sanger.ac.uk/resources/software/lookseq/
A web-based application for alignment visualization, browsing and analysis of genome sequence data.
Proper citation: LookSeq (RRID:SCR_005625) Copy
http://virome.diagcomputing.org/#view=home
A web-application designed for scientific exploration of metagenome sequence data collected from viral assemblages occurring within a number of different environmental contexts. The VIROME informatics pipeline focuses on the classification of predicted open-reading frames (ORFs) from viral metagenomes. The portal allows you to submit your viral metagenome to be processed through the VIROME analysis pipeline, and enable you to investigate your data via the VIROME user interface., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: VIROME (RRID:SCR_004362) Copy
A collaborative ontology for the definition of sequence features used in biological sequence annotation. SO was initially developed by the Gene Ontology Consortium. Contributors to SO include the GMOD community, model organism database groups such as WormBase, FlyBase, Mouse Genome Informatics group, and institutes such as the Sanger Institute and the EBI. Input to SO is welcomed from the sequence annotation community. The OBO revision is available here: http://sourceforge.net/p/song/svn/HEAD/tree/ SO includes different kinds of features which can be located on the sequence. Biological features are those which are defined by their disposition to be involved in a biological process. Biomaterial features are those which are intended for use in an experiment such as aptamer and PCR_product. There are also experimental features which are the result of an experiment. SO also provides a rich set of attributes to describe these features such as polycistronic and maternally imprinted. The Sequence Ontologies use the OBO flat file format specification version 1.2, developed by the Gene Ontology Consortium. The ontology is also available in OWL from Open Biomedical Ontologies. This is updated nightly and may be slightly out of sync with the current obo file. An OWL version of the ontology is also available. The resolvable URI for the current version of SO is http://purl.obolibrary.org/obo/so.owl.
Proper citation: SO (RRID:SCR_004374) Copy
http://metagenomics.atc.tcs.com/binning/ProViDE/
A similarity based binning algorithm that uses a customized set of alignment parameter thresholds / ranges, specifically suited for the accurate taxonomic labelling of viral metagenomic sequences., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: ProViDE (RRID:SCR_004709) Copy
http://caintegrator-info.nci.nih.gov/rembrandt
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on April 28,2023. REMBRANDT is a data repository containing diverse types of molecular research and clinical trials data related to brain cancers, including gliomas, along with a wide variety of web-based analysis tools that readily facilitate the understanding of critical correlations among the different data types. REMBRANDT aims to be the access portal for a national molecular, genetic, and clinical database of several thousand primary brain tumors that is fully open and accessible to all investigators (including intramural and extramural researchers), as well as the public at-large. The main focus is to molecularly characterize a large number of adult and pediatric primary brain tumors and to correlate those data with extensive retrospective and prospective clinical data. Specific data types hosted here are gene expression profiles, real time PCR assays, CGH and SNP array information, sequencing data, tissue array results and images, proteomic profiles, and patients'''' response to various treatments. Clinical trials'''' information and protocols are also accessible. The data can be downloaded as raw files containing all the information gathered through the primary experiments or can be mined using the informatics support provided. This comprehensive brain tumor data portal will allow for easy ad hoc querying across multiple domains, thus allowing physician-scientists to make the right decisions during patient treatments., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Repository of molecular brain neoplasia data (RRID:SCR_004704) Copy
THIS RESOURCE IS NO LONGER IN SERVICE, documented Jan 13, 2022; To enhance the understanding of the evolution of the Kingdom Fungi, 1500+ species were sampled for eight gene loci across all major fungal clades, plus a subset of taxa for a suite of morphological and ultrastructural characters with resulting data: AFTOL Molecular Database (generated by WASABI - Web Accessible Sequence Analysis for Biological Inference), Blast search the AFTOL Database (generated by WASABI), AFTOL primers (generated by WASABI), AFTOL primers by species (generated by WASABI), AFTOL alignments, and the AFTOL Structural and Biochemical Database. Users may submit samples to the AFTOL project. AFTOL is a collaboration centered around four universities in the United States: Duke University (Francois Lutzoni and Rytas Vilgalys), Clark University (David Hibbett), Oregon State University (Joey Spatafora), and University of Minnesota (David McLaughlin). Participants throughout the world have donated vouchers, taxon samples, and gene sequences. The aim of the project is to reconstruct the fungal tree of life using all available data for eight loci (nuclear ribosomal DNA: LSU, SSU, ITS (including 5.8s, ITS1 and ITS2); RNA polymerase II: RPB1, RPB2; elongation factor 1-alpha; mitochondrial SSU rDNA, and mitochondrial ATP synthase protein subunit 6). A further objective of this study is to summarize and integrate current knowledge regarding fungal subcellular features within this new phylogenetic framework. The name of the bioinformatic package developed for AFTOL is WASABI which provides an efficient communication platform to facilitate the collection and dissemination of molecular data to (and from) the laboratories and participants. All molecular data can be viewed, downloaded, verified, and corrected by the participants of AFTOL. A central goal of the WASABI interface is to establish an automated analysis framework that includes basecalling of newly generated chromatograms, contig assembly, quality verification of sequences (including a local BLAST), sequence alignment, and congruence test. Gene sequences that pass all tests and are finally verified by their authors will undergo automated phylogenetic analysis on a regular schedule. Although all steps are initially carried out noninteractively, the users can verify and correct the results at any step and thus initiate the reanalysis of dependent data.
Proper citation: AFTOL (RRID:SCR_004650) Copy
https://github.com/uclinfectionimmunity/Decombinator
Software suite for analysis of T cell receptor repertoire data. Used for fast, efficient analysis of T cell receptor (TcR) repertoire samples, designed to be accessible to those with no previous programming experience.
Proper citation: Decombinator (RRID:SCR_006732) Copy
Database of peer-reviewed, continually updated annotation for the Pseudomonas aeruginosa PAO1 reference strain genome expanded to include all Pseudomonas species to facilitate cross-strain and cross-species genome comparisons with high quality comparative genomics. The database contains robust assessment of orthologs, a novel ortholog clustering method, and incorporates five views of the data at the sequence and annotation levels (Gbrowse, Mauve and custom views) to facilitate genome comparisons. Other features include more accurate protein subcellular localization predictions and a user-friendly, Boolean searchable log file of updates for the reference strain PAO1. The current annotation is updated using recent research literature and peer-reviewed submissions by a worldwide community of PseudoCAP (Pseudomonas aeruginosa Community Annotation Project) participating researchers. If you are interested in participating, you are invited to get involved. Many annotations, DNA sequences, Orthologs, Intergenic DNA, and Protein sequences are available for download.
Proper citation: Pseudomonas Genome Database (RRID:SCR_006590) Copy
Collection of data related to crop plant and model organism Zea mays. Used to synthesize, display, and provide access to maize genomics and genetics data, prioritizing mutant and phenotype data and tools, structural and genetic map sets, and gene models and to provide support services to the community of maize researchers. Data stored at MaizeGDB was inherited from the MaizeDB and ZmDB projects. Sequence data are from GenBank. Data are searchable by phenotype, traits, Pests, Gel Pattern, and Mutant Images.
Proper citation: MaizeGDB (RRID:SCR_006600) Copy
http://sourceforge.net/projects/gasic/
A method to correct read alignment results for the ambiguities imposed by similarities of genomes.
Proper citation: GASiC (RRID:SCR_006765) Copy
International collaboration producing an extensive public catalog of human genetic variation, including SNPs and structural variants, and their haplotype contexts, in an effort to provide a foundation for investigating the relationship between genotype and phenotype. The genomes of about 2500 unidentified people from about 25 populations around the world were sequenced using next-generation sequencing technologies. Redundant sequencing on various platforms and by different groups of scientists of the same samples can be compared. The results of the study are freely and publicly accessible to researchers worldwide. The consortium identified the following populations whose DNA will be sequenced: Yoruba in Ibadan, Nigeria; Japanese in Tokyo; Chinese in Beijing; Utah residents with ancestry from northern and western Europe; Luhya in Webuye, Kenya; Maasai in Kinyawa, Kenya; Toscani in Italy; Gujarati Indians in Houston; Chinese in metropolitan Denver; people of Mexican ancestry in Los Angeles; and people of African ancestry in the southwestern United States. The goal Project is to find most genetic variants that have frequencies of at least 1% in the populations studied. Sequencing is still too expensive to deeply sequence the many samples being studied for this project. However, any particular region of the genome generally contains a limited number of haplotypes. Data can be combined across many samples to allow efficient detection of most of the variants in a region. The Project currently plans to sequence each sample to about 4X coverage; at this depth sequencing cannot provide the complete genotype of each sample, but should allow the detection of most variants with frequencies as low as 1%. Combining the data from 2500 samples should allow highly accurate estimation (imputation) of the variants and genotypes for each sample that were not seen directly by the light sequencing. All samples from the 1000 genomes are available as lymphoblastoid cell lines (LCLs) and LCL derived DNA from the Coriell Cell Repository as part of the NHGRI Catalog. The sequence and alignment data generated by the 1000genomes project is made available as quickly as possible via their mirrored ftp sites. ftp://ftp.1000genomes.ebi.ac.uk ftp://ftp-trace.ncbi.nlm.nih.gov/1000genomes
Proper citation: 1000 Genomes: A Deep Catalog of Human Genetic Variation (RRID:SCR_006828) Copy
http://weizhong-lab.ucsd.edu/cd-hit-otu/
Data analysis service and software program that perform Operantional Taxonomic Units (OTUs) finding. It uses a three-step clustering for identifying OTUs. The first-step clustering is raw read filtering and trimming. The second step is error-free reads picking.. At the last step, OTU clustering is done at different distanct cutoffs (0.01, 0.02, 0.03... 0.12).
Proper citation: CD-HIT-OTU (RRID:SCR_006983) Copy
The HumanCyc database describes human metabolic pathways and the human genome. By presenting metabolic pathways as an organizing framework for the human genome, HumanCyc provides the user with an extended dimension for functional analysis of Homo sapiens at the genomic level. A computational pathway analysis of the human genome assigned human enzymes to predicted metabolic pathways. Pathway assignments place genes in their larger biological context, and are a necessary step toward quantitative modeling of metabolism. HumanCyc contains the complete genome sequence of Homo sapiens, as presented in Build 31. Data on the human genome from Ensembl, LocusLink and GenBank were carefully merged to create a minimally redundant human gene set to serve as an input to SRI''s PathoLogic software, which generated the database and predicted Homo sapiens metabolic pathways from functional information contained in the genome''s annotation. SRI did not re-annotate the genome, but worked with the gene function assignments in Ensembl, LocusLink, and GenBank. The resulting pathway/genome database (PGDB) includes information on 28,783 genes, their products and the metabolic reactions and pathways they catalyze. Also included are many links to other databases and publications. The Pathway Tools software/database bundle includes HumanCyc and the Pathway Tools software suite and is available under license. This form of HumanCyc is faster and more powerful than the Web version.
Proper citation: HumanCyc: Encyclopedia of Homo sapiens Genes and Metabolism (RRID:SCR_007050) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the PRECISE-TBI Resources search. From here you can search through a compilation of resources used by PRECISE-TBI and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that PRECISE-TBI has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on PRECISE-TBI then you can log in from here to get additional features in PRECISE-TBI such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into PRECISE-TBI you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within PRECISE-TBI that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.