Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
https://fungi.ensembl.org/Neurospora_crassa/Info/Index
It's strategy involves Whole Genome Shotgun (WGS) sequencing, in which sequence from the entire genome is generated and reassembled. This method is standard for microbial genome sequencing, and has been successfully applied to Drosophila. Neurospora is an ideal candidate for this approach because of the low repeat content of the genome. Neurospora crassa Database has expanded the scope of its database by including a mitochondrial annotation, incorporating information from the Neurospora compendium, and assigning NCU numbers to tRNA and rRNAs. They have improved the annotation process to predict untranslated regions and to reduce the number of spurious predictions. As a result, version 3 contains 9,826 genes, 794 fewer than version 2. During the initial phase of a WGS project they sequence both ends of the 4 kb inserts from a plasmid library prepared using randomly sheared and sized-selected DNA. The shotgun reads are assembled by recognizing overlapping regions of sequence and making use of the knowledge of the orientation and distance of the paired reads from each plasmid. Obtaining deep sequence coverage though high levels of sequence redundancy assures that the majority of the genome is represented in the initial assembly and that the consensus sequence is of high quality. Their approach toward the initial assembly was conservative, meaning they would rather fail to join sequence contigs that might overlap each other than risk making false joins between two closely related but non-overlapping genomic regions. Hence, the initial assembly contains many sequence contigs and over time these contigs will increase in size and decrease in number as they are joined together. After shotgun sequencing and assembly there was a second phase of sequencing in which additional sequence was obtained from specific regions that were missing from the original assembly or are recognized to be of low quality in the consensus. The Neurospora crassa sequencing project reflects a close collaboration between the Broad Institute and the Neurospora research community. Principal investigators include Bruce Birren and Chad Nusbaum from the Broad Institute, Matt Sachs at the Oregon Graduate Institute of Science and Technology, Chuck Staben at the University of Kentucky and Jak Kinsey at the Fungal Genetics Stock Center at the University of Kansas Medical Center. In addition, we have a larger Advisory Board made up of a number of Neurospora researchers. Sponsors: They have been funded by the National Science Foundation to sequence the N. crassa genome and make the information publicly available.
Proper citation: Neurospora crassa Database (RRID:SCR_001372) Copy
Database and browser that provides a central resource to archive and display association between genetic variation and high-throughput molecular-level phenotypes. This effort originated with the NIH GTEx roadmap project: however the scope of this resource will be extended to include any available genotype/molecular phenotype datasets.
Proper citation: GTEx eQTL Browser (RRID:SCR_001618) Copy
http://www.ncbi.nlm.nih.gov/cdd
Database of annotations of functional units in proteins including multiple sequence alignment models for ancient domains and full-length proteins. This collection of models includes 3D structures that display the sequence/structure/function relationships in proteins. It also includes alignments of the domains to known three-dimensional protein structures in the MMDB database. The source databases are Pfam, Smart, and COG. Users can identify amino acids in protein sequences with the resources available as well as view single sequences embedded within multiple sequence alignments.
Proper citation: Conserved Domain Database (RRID:SCR_002077) Copy
http://www.cbs.dtu.dk/services/SignalP/
Web application for prediction of the presence and location of signal peptide cleavage sites in amino acid sequences from different organisms. The method incorporates a prediction of cleavage sites and a signal peptide/non-signal peptide prediction based on a combination of several artificial neural networks.
Proper citation: SignalP (RRID:SCR_015644) Copy
http://dynamine.ibsquare.be/submission/
An NMR based method for protein folding prediction. Users can enter a UniProt identifier, FASTA sequences, or upload a file containing FASTA sequences and results are returned., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: DynaMine (RRID:SCR_014559) Copy
A resource for information pertaining to methodologies, tools and technologies of gene expression. The website offers resources for sequence analysis, database services, and other technologies of gene expression and regulation.
Proper citation: IFTI-Mirage (RRID:SCR_000505) Copy
http://sourceforge.net/projects/skewer/
Software program for adapter trimming that is specially designed for processing Illumina paired-end sequences.
Proper citation: skewer (RRID:SCR_001151) Copy
Consortium represents all publicly available gene trap cell lines, which are available on non-collaborative basis for nominal handling fees. Researchers can search and browse IGTC database for cell lines of interest using accession numbers or IDs, keywords, sequence data, tissue expression profiles and biological pathways, can find trapped genes of interest on IGTC website, and order cell lines for generation of mutant mice through blastocyst injection. Consortium members include: BayGenomics (USA), Centre for Modelling Human Disease (Toronto, Canada), Embryonic Stem Cell Database (University of Manitoba, Canada), Exchangeable Gene Trap Clones (Kumamoto University, Japan), German Gene Trap Consortium provider (Germany), Sanger Institute Gene Trap Resource (Cambridge, UK), Soriano Lab Gene Trap Resource (Mount Sinai School of Medicine, New York, USA), Texas Institute for Genomic Medicine - TIGM (USA), TIGEM-IRBM Gene Trap (Naples, Italy).
Proper citation: International Gene Trap Consortium (RRID:SCR_002305) Copy
A collection of high quality multiple sequence alignments for objective, comparative studies of alignment algorithms. The alignments are constructed based on 3D structure superposition and manually refined to ensure alignment of important functional residues. A number of subsets are defined covering many of the most important problems encountered when aligning real sets of proteins. It is specifically designed to serve as an evaluation resource to address all the problems encountered when aligning complete sequences. The first release provided sets of reference alignments dealing with the problems of high variability, unequal repartition and large N/C-terminal extensions and internal insertions. Version 2.0 of the database incorporates three new reference sets of alignments containing structural repeats, trans-membrane sequences and circular permutations to evaluate the accuracy of detection/prediction and alignment of these complex sequences.
Within the resource, users can look at a list of all the alignments, download the whole database by ftp, get the "c" program to compare a test alignment with the BAliBASE reference (The source code for the program is freely available), or look at the results of a comparison study of several multiple alignment programs, using BAliBASE reference sets.
Proper citation: BAliBASE (RRID:SCR_001940) Copy
Issue
Software package for analysis of brain imaging data sequences. Sequences can be a series of images from different cohorts, or time-series from same subject. Current release is designed for analysis of fMRI, PET, SPECT, EEG and MEG.
Proper citation: SPM (RRID:SCR_007037) Copy
https://www.ncbi.nlm.nih.gov/sutils/pasc/viridty.cgi
Web tool for analysis of pairwise identity distribution within viral families. Used for virus sequence-based classification. Data in the system are updated every day to reflect changes in virus taxonomy and additions of new virus sequences to the public database.
Proper citation: PASC (RRID:SCR_016642) Copy
https://github.com/schloi/MARVEL
Software set of tools that facilitate overlapping, patching, correction and assembly of noisy long reads.
Proper citation: Marvel (RRID:SCR_017621) Copy
https://github.com/davidemms/OrthoFinder
Software Python application for comparative genomics analysis. Finds orthogroups and orthologs, infers rooted gene trees for all orthogroups and identifies all of gene duplcation events in those gene trees, infers rooted species tree for species being analysed and maps gene duplication events from gene trees to branches in species tree, improves orthogroup inference accuracy. Runs set of protein sequence files, one per species, in FASTA format.
Proper citation: OrthoFinder (RRID:SCR_017118) Copy
http://www.imgt.org/HighV-QUEST/home.action
Next generation B and T cell sequence alignment and characterization online surface by IMGT. Web portal for immunoglobulin (IG) or antibody and T cell receptor (TR) analysis from NGS high throughput and deep sequencing.
Proper citation: IMGT HighV-QUEST (RRID:SCR_018196) Copy
http://amphoranet.pitgroup.org/
Webserver implementation of the AMPHORA2 workflow for phylogenetic analysis of metagenomic shotgun sequencing data. It is capable of assigning a probability-weighted taxonomic group for each phylogenetic marker gene found in the input metagenomic sample.
Proper citation: AmphoraNet (RRID:SCR_005009) Copy
http://www.ebi.ac.uk/biosamples/
Database that aggregates sample information for reference samples (e.g. Coriell Cell lines) and samples for which data exist in one of the EBI''''s assay databases such as ArrayExpress, the European Nucleotide Archive or PRoteomics Identificates DatabasE. It provides links to assays for specific samples, and accepts direct submissions of sample information. The goals of the BioSample Database include: # recording and linking of sample information consistently within EBI databases such as ENA, ArrayExpress and PRIDE; # minimizing data entry efforts for EBI database submitters by enabling submitting sample descriptions once and referencing them later in data submissions to assay databases and # supporting cross database queries by sample characteristics. The database includes a growing set of reference samples, such as cell lines, which are repeatedly used in experiments and can be easily referenced from any database by their accession numbers. Accession numbers for the reference samples will be exchanged with a similar database at NCBI. The samples in the database can be queried by their attributes, such as sample types, disease names or sample providers. A simple tab-delimited format facilitates submissions of sample information to the database, initially via email to biosamples (at) ebi.ac.uk. Current data sources: * European Nucleotide Archive (424,811 samples) * PRIDE (17,001 samples) * ArrayExpress (1,187,884 samples) * ENCODE cell lines (119 samples) * CORIELL cell lines (27,002 samples) * Thousand Genome (2,628 samples) * HapMap (1,417 samples) * IMSR (248,660 samples)
Proper citation: BioSample Database at EBI (RRID:SCR_004856) Copy
A database of protein families, each represented by multiple sequence alignments and hidden Markov models (HMMs). Users can analyze protein sequences for Pfam matches, view Pfam family annotation and alignments, see groups of related families, look at the domain organization of a protein sequence, find the domains on a PDB structure, and query Pfam by keywords. There are two components to Pfam: Pfam-A and Pfam-B. Pfam-A entries are high quality, manually curated families that may automatically generate a supplement using the ADDA database. These automatically generated entries are called Pfam-B. Although of lower quality, Pfam-B families can be useful for identifying functionally conserved regions when no Pfam-A entries are found. Pfam also generates higher-level groupings of related families, known as clans (collections of Pfam-A entries which are related by similarity of sequence, structure or profile-HMM).
Proper citation: Pfam (RRID:SCR_004726) Copy
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 11,2023. SuperCAT hosts typing databases for the Bacillus cereus group of bacteria. The databases contain MultiLocus Sequence Typing (MLST), MultiLocus Enzyme Electrophoresis (MLEE), and Amplified Fragment Length Polymorphism (AFLP) phylogenetic data. multilocus, sequence, Bacillus cereus, bacteria, Genomics, non-vertebrate, taxonomy, identification
Proper citation: SuperCAT (RRID:SCR_004882) Copy
A clade oriented, community curated database containing genomic, genetic, phenotypic and taxonomic information for plant genomes. Genomic information is presented in a comparative format and tied to important plant model species such as Arabidopsis. SGN provides tools such as: BLAST searches, the SolCyc biochemical pathways database, a CAPS experiment designer, an intron detection tool, an advanced Alignment Analyzer, and a browser for phylogenetic trees. The SGN code and database are developed as an open source project, and is based on database schemas developed by the GMOD project and SGN-specific extensions.
Proper citation: SGN (RRID:SCR_004933) Copy
http://cgi-www.daimi.au.dk/cgi-chili/datfap/frontdoor.py
A database of transcription factors from 13 plant species, and PCR primers for around 90% of them.
Proper citation: DATFAP (RRID:SCR_005413) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the ASWG Resources search. From here you can search through a compilation of resources used by ASWG and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that ASWG has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on ASWG then you can log in from here to get additional features in ASWG such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into ASWG you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within ASWG that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.