Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
http://users-birc.au.dk/biopv/php/fabox/
Tools for splitting, joining and otherwise manipulating FASTA format sequence files. The first tools in the toolbox is for manipulating fasta headers, cropping alignments and doing some sequence comparison allowing users to combine the description of data (often in excel spreadsheets) with the actual data (often DNA sequences). Also, producing correct input files for a range of programs seems to be problematic for the average user. Hence, some converters in some of the services have been included as well as some stand-alone converters. The converters are not necessarily meant to provide the final input file, but you''ll get a valid input file for Arlequin, MrBayes etc. - that you may further edit so it suit your needs. This means that you may need to combine several of the tools to finish your handling - but it keeps it relatively simple to use. Please note that FaBox is written in PHP and ONLY RUNS ON A WEBSERVER.
Proper citation: FaBox (RRID:SCR_005350) Copy
http://mesquiteproject.org/packages/chromaseq/
A software package in Mesquite that processes chromatograms, makes contigs, base calls, etc., using in part the programs Phred and Phrap.
Proper citation: Chromaseq (RRID:SCR_005587) Copy
http://nucleobytes.com/index.php/4peaks
Software application for viewing and editing sequence trace files.
Proper citation: 4Peaks (RRID:SCR_000015) Copy
NIH initiative to support production of cDNA libraries, clones and 5'/3' sequences and to provide set of full-length (open reading frame) sequences and cDNA clones of expressed genes for Xenopus laevis and Xenopus tropicalis. Clones distribution is outsourced to for profit companies. Project concluded in September 2008. Resources generated by XGC are publicly accessible to biomedical research community. All sequences are deposited into GenBank.Corresponding clones are available through IMAGE clone distribution network. With conclusion of XGC project, GenBank records of XGC sequences will be frozen, without further updates. Since knowledge of what constitutes full-length coding region for some of genes and transcripts for which we have XGC clones will likely change in future, users planning to order XGC clones will need to monitor for these changes. Users can make use of genome browsers and gene-specific databases, such as UCSC Genome browser, NCBI's Map Viewer, and Entrez Gene, to view relevant regions of genome (browsers) or gene-related information (Entrez Gene).
Proper citation: Xenopus Gene Collection (RRID:SCR_007023) Copy
Part of zebrafish genome project. ZGC project to produce cDNA libraries, clones and sequences to provide complete set of full-length (open reading frame) sequences and cDNA clones of expressed genes for zebrafish. All ZGC sequences are deposited in GenBank and clones can be purchased from distributors of IMAGE consortium. With conclusion of ZGC project in September 2008, GenBank records of ZGC sequences will be frozen, without further updates. Since definition of what constitutes full-length coding region for some of genes and transcripts for which we have ZGC clones will likely change in future, users planning to order ZGC clones will need to monitor for these changes. Users can make use of genome browsers and gene-specific databases, such as UCSC Genome browser, NCBI's Map Viewer, and Entrez Gene, to view relevant regions of genome (browsers) or gene-related information (Entrez Gene).
Proper citation: Zebrafish Gene Collection (RRID:SCR_007054) Copy
https://github.com/ekg/fastahack
Software application for indexing and extracting sequences and subsequences from FASTA files. It will only generate indexes for FASTA files in which the sequences have self-consistent line lengths.
Proper citation: Fastahack (RRID:SCR_016090) Copy
Consortium represents all publicly available gene trap cell lines, which are available on non-collaborative basis for nominal handling fees. Researchers can search and browse IGTC database for cell lines of interest using accession numbers or IDs, keywords, sequence data, tissue expression profiles and biological pathways, can find trapped genes of interest on IGTC website, and order cell lines for generation of mutant mice through blastocyst injection. Consortium members include: BayGenomics (USA), Centre for Modelling Human Disease (Toronto, Canada), Embryonic Stem Cell Database (University of Manitoba, Canada), Exchangeable Gene Trap Clones (Kumamoto University, Japan), German Gene Trap Consortium provider (Germany), Sanger Institute Gene Trap Resource (Cambridge, UK), Soriano Lab Gene Trap Resource (Mount Sinai School of Medicine, New York, USA), Texas Institute for Genomic Medicine - TIGM (USA), TIGEM-IRBM Gene Trap (Naples, Italy).
Proper citation: International Gene Trap Consortium (RRID:SCR_002305) Copy
https://bionanogenomics.com/wp-content/uploads/2017/01/30047-Irys-User-Guide.pdf
System by BioNano Genomics ( formerly BioNanomatrix) which provides optical next generation mapping (NGM). Used for sequence assembly and structural variation analysis. Provides Scaffold Bionano genome mapping data with sequencing data to improve assembly contiguity, reduce sequencing coverage needed, and automatically correct errors in sequencing based assemblies.
Proper citation: BioNano: Irys system (RRID:SCR_016754) Copy
http://www.open-ephys.org/pulsepal
Open source pulse train generator that allows users to create and trigger software defined trains of voltage pulses with high temporal precision. Generates precisely timed pulse sequences for use in research involving electrophysiology or psychophysics.
Proper citation: Open Ephys: Pulse Pal (RRID:SCR_017203) Copy
A collection of high quality multiple sequence alignments for objective, comparative studies of alignment algorithms. The alignments are constructed based on 3D structure superposition and manually refined to ensure alignment of important functional residues. A number of subsets are defined covering many of the most important problems encountered when aligning real sets of proteins. It is specifically designed to serve as an evaluation resource to address all the problems encountered when aligning complete sequences. The first release provided sets of reference alignments dealing with the problems of high variability, unequal repartition and large N/C-terminal extensions and internal insertions. Version 2.0 of the database incorporates three new reference sets of alignments containing structural repeats, trans-membrane sequences and circular permutations to evaluate the accuracy of detection/prediction and alignment of these complex sequences.
Within the resource, users can look at a list of all the alignments, download the whole database by ftp, get the "c" program to compare a test alignment with the BAliBASE reference (The source code for the program is freely available), or look at the results of a comparison study of several multiple alignment programs, using BAliBASE reference sets.
Proper citation: BAliBASE (RRID:SCR_001940) Copy
http://bioinformatics.udel.edu/Research/skategenomeproject
Core facility provides a model for collaborative approaches to use specialized resources and expertise in an integrated process. Core builds on the expertise and resources provided by the Bioinformatics Cores of the five northeastern states that form NECC. The Skate Genome Annotation Workshops and Jamborees offer training and opportunities for faculty and students to work with and annotate genome sequences. Workshops include lectures, tutorials and exercises annotating the genome of the little skate, Leucoraja erinacea.
Proper citation: University of Delaware Skate Genome Project (RRID:SCR_005300) Copy
http://linux1.softberry.com/spldb/SpliceDB.html
Database of canonical and non-canonical mammalian splice sites. The information about verified splice site sequences for canonical and non-canonical sites is presented with the supporting evidence. Weight matrices were built for the major splice groups, which can be incorporated into gene prediction programs.
Proper citation: SpliceDB (RRID:SCR_006262) Copy
http://www.sanger.ac.uk/Projects/Fungi/
Fungal genomes available from the Sanger Institute. Data are accessible in a number of ways; for each organism there is a BLAST server, allowing search of the sequences. Sequences can also be down-loaded directly by FTP. In addition, for those organisms being sequenced using a cosmid approach, finished and annotated cosmids are submitted to EMBL and other public databases.
Proper citation: Fungi Sequencing Projects (RRID:SCR_008524) Copy
http://depts.washington.edu/yeastrc/
Biomedical technology research center that (1) exploits the budding yeast Saccharomyces cerevisiae to develop novel technologies for investigating and characterizing protein function and protein structure (2) facilitates research and extension of new technologies through collaboration, and (3) actively disseminates data and technology to the research community. Through collaboration, the YRC freely provides resources and expertise in six core technology areas: Protein Tandem Mass Spectrometry, Protein Sequence-Function Relationships, Quantitative Phenotyping, Protein Structure Prediction and Design, Fluorescence Microscopy, Computational Biology.
Proper citation: Yeast Resource Center (RRID:SCR_007942) Copy
http://www.sanger.ac.uk/science/tools/ssaha2-0
A program designed for the efficient mapping of sequence reads onto genomic references. The software is capable of reading most sequencing platforms and giving a range of outputs are supported.
Proper citation: Sequence Search and Alignment by Hashing Algorithm (RRID:SCR_000544) Copy
The EBI genomes pages give access to a large number of complete genomes including bacteria, archaea, viruses, phages, plasmids, viroids and eukaryotes. Methods using whole genome shotgun data are used to gain a large amount of genome coverage for an organism. WGS data for a growing number of organisms are being submitted to DDBJ/EMBL/GenBank. Genome entries have been listed in their appropriate category which may be browsed using the website navigation tool bar on the left. While organelles are all listed in a separate category, any from Eukaryota with chromosome entries are also listed in the Eukaryota page. Within each page, entries are grouped and sorted at the species level with links to the taxonomy page for that species separating each group. Within each species, entries whose source organism has been categorized further are grouped and numbered accordingly. Links are made to: * taxonomy * complete EMBL flatfile * CON files * lists of CON segments * Project * Proteomes pages * FASTA file of Proteins * list of Proteins
Proper citation: EBI Genomes (RRID:SCR_002426) Copy
http://www.broad.mit.edu/annotation/fungi/fgi/
Produces and analyzes sequence data from fungal organisms that are important to medicine, agriculture and industry. The FGI is a partnership between the Broad Institute and the wider fungal research community, with the selection of target genomes governed by a steering committee of fungal scientists. Organisms are selected for sequencing as part of a cohesive strategy that considers the value of data from each organism, given their role in basic research, health, agriculture and industry, as well as their value in comparative genomics.
Proper citation: Fungal Genome Initiative (RRID:SCR_003169) Copy
A fungal rDNA internal transcribed spacer (ITS) sequence database (although additional genes and genetic markers are also welcome) to facilitate identification of environmental samples of fungal DNA. Additional important features include user annotation of INSD sequences to add metadata on, e.g., locality, habitat, soil, climate, and interacting taxa. The user can furthermore annotate INSD sequences with additional species identifications that will appear in the results of any analyses done. UNITE focuses on high-quality ITS sequences generated from fruiting bodies collected and identified by experts and deposited in public herbaria. In addition, it also holds all fungal ITS sequences in the International Nucleotide Sequence Databases (INSD: NCBI, EMBL, DDBJ). Both sets of sequences may be used in any analyses carried out. UNITE is accompanied by a project management system called PlutoF, where users can store field data, document the sequencing lab procedures, manage sequences, and make analyses. PlutoF intends to make it possible for taxonomists, ecologists, and biogeographers to use a common platform for data storage, handling, and analyses, with the intent of facilitating an integration of these disciplines. A user can have an unlimited number of projects but still make analyses across any project data available to him.
Proper citation: UNITE (RRID:SCR_006518) Copy
A comparative platform for green plant genomics. Families of orthologous and paralogous genes that represent the modern descendents of ancestral gene sets are constructed at key phylogenetic nodes. These families allow easy access to clade specific orthology / paralogy relationships as well as clade specific genes and gene expansions. As of release v9.1, Phytozome provides access to forty-one sequenced and annotated green plant genomes which have been clustered into gene families at 20 evolutionarily significant nodes. Where possible, each gene has been annotated with PFAM, KOG, KEGG, and PANTHER assignments, and publicly available annotations from RefSeq, UniProt, TAIR, JGI are hyper-linked and searchable., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025.
Proper citation: Phytozome (RRID:SCR_006507) Copy
http://www.ncbi.nlm.nih.gov/projects/genome/assembly/grc/
Consortium that puts sequences into a chromosome context and provides the best possible reference assembly for human, mouse, and zebrafish via FTP. Tools to facilitate the curation of genome assemblies based on the sequence overlaps of long, high quality sequences.
Proper citation: Genome Reference Consortium (RRID:SCR_006553) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the ASWG Resources search. From here you can search through a compilation of resources used by ASWG and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that ASWG has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on ASWG then you can log in from here to get additional features in ASWG such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into ASWG you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within ASWG that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.