Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.
https://github.com/DerrickWood/kraken2
Software tool as second version of Kraken taxonomic sequence classification system.
Proper citation: kraken2 (RRID:SCR_026838) Copy
http://trace.ddbj.nig.ac.jp/dor/index_e.html
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 6,2023. Archival database of functional genomics data generated by microarray and highly parallel new generation sequencers. Data are exchanged between ArrayExpress at EBI and DOR in common MAGE-TAB format. Supports MIAME and MINSEQE-compliant data submissions. DOR issues accession numbers, E-DORD-n to experiment and A-DORD-n to array design. DOR exchanges public data with the EBI ArrayExpress in common MAGE-TAB format. Note: At present, DOR does not accept submissions. DDBJ will announce launch of DOR when it is ready. (2013/01/31) The data can be kept private until your paper is published. You can set the hold date for a maximum of 1 year and can change it. Registered records are released according to the Data Release Policy.
Proper citation: DDBJ Omics Archive (RRID:SCR_000597) Copy
A resource for information pertaining to methodologies, tools and technologies of gene expression. The website offers resources for sequence analysis, database services, and other technologies of gene expression and regulation.
Proper citation: IFTI-Mirage (RRID:SCR_000505) Copy
Consortium represents all publicly available gene trap cell lines, which are available on non-collaborative basis for nominal handling fees. Researchers can search and browse IGTC database for cell lines of interest using accession numbers or IDs, keywords, sequence data, tissue expression profiles and biological pathways, can find trapped genes of interest on IGTC website, and order cell lines for generation of mutant mice through blastocyst injection. Consortium members include: BayGenomics (USA), Centre for Modelling Human Disease (Toronto, Canada), Embryonic Stem Cell Database (University of Manitoba, Canada), Exchangeable Gene Trap Clones (Kumamoto University, Japan), German Gene Trap Consortium provider (Germany), Sanger Institute Gene Trap Resource (Cambridge, UK), Soriano Lab Gene Trap Resource (Mount Sinai School of Medicine, New York, USA), Texas Institute for Genomic Medicine - TIGM (USA), TIGEM-IRBM Gene Trap (Naples, Italy).
Proper citation: International Gene Trap Consortium (RRID:SCR_002305) Copy
A collection of high quality multiple sequence alignments for objective, comparative studies of alignment algorithms. The alignments are constructed based on 3D structure superposition and manually refined to ensure alignment of important functional residues. A number of subsets are defined covering many of the most important problems encountered when aligning real sets of proteins. It is specifically designed to serve as an evaluation resource to address all the problems encountered when aligning complete sequences. The first release provided sets of reference alignments dealing with the problems of high variability, unequal repartition and large N/C-terminal extensions and internal insertions. Version 2.0 of the database incorporates three new reference sets of alignments containing structural repeats, trans-membrane sequences and circular permutations to evaluate the accuracy of detection/prediction and alignment of these complex sequences.
Within the resource, users can look at a list of all the alignments, download the whole database by ftp, get the "c" program to compare a test alignment with the BAliBASE reference (The source code for the program is freely available), or look at the results of a comparison study of several multiple alignment programs, using BAliBASE reference sets.
Proper citation: BAliBASE (RRID:SCR_001940) Copy
http://icebox.lbl.gov:8080/ApolloWebDemo/jbrowse/
WebApollo is an extensible web-based sequence annotation editor for community annotation. No software download is required and the annotations are saved to a centralized database with real-time annotation updating. (The edit server mediates annotation changes made by multiple users.) The Web based client uses JBrowse, is fast and highly interactive. WebApollo accesses many types of genomic data including access to public data from UCSC, Ensembl, and GMOD Chado databases. Source code (BSD License) * Client source code: https://github.com/berkeleybop/jbrowse * Annotation editing engine: http://code.google.com/p/apollo-web * Data model and I/O layer: http://code.google.com/p/gbol * Trellis server code: http://code.google.com/p/genomancer
Proper citation: WebApollo: A Web-Based Sequence Annotation Editor for Community Annotation (RRID:SCR_005321) Copy
http://igs-server.cnrs-mrs.fr/mgdb/Rickettsia/
THIS RESOURCE IS NO LONGER IN SERVICE, documented August 18, 2016. Rickettsia are obligate intracellular bacteria living in arthropods. They occasionally cause diseases in humans. To understand their pathogenicity, physiologies and evolutionary mechanisms, RicBase is sequencing different species of Rickettsia. Up to now we have determined the genome sequences of R. conorii, R. felis, R. bellii, R. africae, and R. massiliae. The RicBase aims to organize the genomic data to assist followup studies of Rickettsia. This website contains information on R. conorii and R. prowazekii. A R. conorii and R. prowazekii comparative genome map is also available. Images of genome maps, dendrogram, and sequence alignment allow users to gain a visualization of the diagrams.
Proper citation: Rickettsia Genome Database (RRID:SCR_007102) Copy
http://wwwmgs.bionet.nsc.ru/mgs/programs/panalyst/
WebProAnalyst provides web-accessible analysis for scanning the quantitative structure-activity relationships in protein families. It searches for a sequence region, whose substitutions are correlated with variations in the activities of a homologous protein set, the so-called activity modulating sites. WebProAnalyst allows users to search for the key physicochemical characteristics of the sites that affect the changes in protein activities. It enables the building of multiple linear regression and neural networks models that relate these characteristics to protein activities. WebProAnalyst implements multiple linear regression analysis, back propagation neural networks and the Structure-Activity Correlation/Determination Coefficient (SACC/SADC). A back propagation neural network is implemented as a two-layered network, one layer as input, the other as output (Rumelhart et al, 1986). WebProAnalyst uses alignment of amino acid sequences and data on protein activity (pK, Km, ED50, among others). The input data are the numerical values for the physicochemical characteristics of a site in the multiple alignment given by a slide window. The output data are the predicted activity values. The current version of WebProAnalyst handles a single activity for a single protein. The SACC/SADC may be defined as an estimate of the strongest multiple correlation between the physicochemical characteristics of a site in a multiple alignment and protein activities. The SACC/SADC coefficient makes possible the calculation of the possible highest correlation achievable for the quantitative relationship between the physicochemical properties of sites and protein activities. The SACC/SADC is a convenient means for an arrangement of positions by their functional significance. WebProAnalyst outputs a list of multiple alignment positions, the respective correlation values, also regression analysis parameters for the relationships between the amino acid physicochemical characteristics at these positions and the protein activity values.
Proper citation: Webproanalyst (RRID:SCR_008348) Copy
http://phylopythias.bifo.helmholtz-hzi.de/index.php?phase=wait
Web Server for Taxonomic Assignment of Metagenome Sequences that is a fast and accurate sequence composition-based classifier that utilizes the hierarchical relationships between clades. Taxonomic assignments with the web server can be made with a generic model, or with sample-specific models that users can specify and create. Several interactive visualization modes and multiple download formats allow quick and convenient analysis and downstream processing of taxonomic assignments.
Proper citation: PhyloPythiaS (RRID:SCR_011923) Copy
Software tool as catalog of inferred sequence binding preferences. Online library of transcription factors and their DNA binding motifs.
Proper citation: CIS-BP (RRID:SCR_017236) Copy
http://cbcsrv.watson.ibm.com/phylopythia.html
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 1, 2023. Data analysis service for accurate phylogenetic classification of variable-length DNA fragments.
Proper citation: PhyloPythia (RRID:SCR_000540) Copy
http://amphoranet.pitgroup.org/
Webserver implementation of the AMPHORA2 workflow for phylogenetic analysis of metagenomic shotgun sequencing data. It is capable of assigning a probability-weighted taxonomic group for each phylogenetic marker gene found in the input metagenomic sample.
Proper citation: AmphoraNet (RRID:SCR_005009) Copy
http://gladyshevlab.org/SelenoproteinPredictionServer/
Web server to predict eukaryotic selenoproteins and SECIS (SElenoCysteine Insertion Sequences) elements along nucleotide sequences. SECISearch3 replaces its predecessor SECISearch as a tool for prediction of eukaryotic SECIS elements. Seblastian is a method for selenoprotein gene detection that uses SECISearch3 and then predicts selenoprotein sequences encoded upstream of SECIS elements. Seblastian is able to both identify known selenoproteins and predict new selenoproteins.
Proper citation: SECISearch3 and Seblastian (RRID:SCR_003186) Copy
A web program that can locate residue periodicities in either amino acid or DNA sequences. It is based on an algorithm of Dr. A.D. McLachlan (1977). NOTE: You must use a Java compatible browser to run the application.
Proper citation: FT (RRID:SCR_006228) Copy
http://athina.biol.uoa.gr/bioinformatics/NON-RED/index.html
A web tool to select biological sequences from a given set, with similarity / homology less than a user-defined level. This web-based application takes as input a set of N sequences and outputs a set of sequences of user-determined redundancy. Initially, the algorithm runs an all-against-all BLAST alignment on the input data set and creates an NxN matrix of pairwise distances defined by the similarity percentages. In the next step, the algorithm removes the sequence with the largest number of neighbors, causing that sequence not to be counted as a neighbor of any other sequence during the next iterations. It then reassesses the number of neighbors of each sequence and repeats the previous step until the sequences left over have no more neighbors. The user can specify the similarity (%) threshold and the minimum coverage length of the alignments. Sequences with a similarity below the threshold or a smaller coverage than the minimum length are not considered to be neighbors.
Proper citation: NON-RED (RRID:SCR_006225) Copy
http://athina.biol.uoa.gr/bioinformatics/waveTM/
A web tool for the prediction of transmembrane segments in alpha-helical membrane proteins. A sliding window of 20 residues is used in order to calculate an average residue hydrophobicity profile, using a hydrophobicity scale. Discrete Wavelet Transform is applied on the average residue hydrophobicity signal and the different frequency coefficients produced are adaptively thresholded so that a denoised signal is reconstructed. A dynamic programming algorithm processes the denoised signal to provide the optimal model for the number, the length and the location of membrane-spanning segments. The end points of the predicted segments are extended to include flanking hydrophobic residues. Topology prediction can also be obtained in conjunction with OrienTM (Liakopoulos et al, 2001). Analysis of a non-redundant test set, provides a ~95% per segment accuracy and ~90% per residue accuracy. Now, you can: * Run waveTM on a sequence * Browse the results obtained with the algorithm * View additional material concerning the hydrophobicity scale
Proper citation: waveTM (RRID:SCR_006199) Copy
http://athina.biol.uoa.gr/PRED-TMR/
A web server that predicts transmembrane domains in proteins using solely information contained in the sequence itself. The algorithm refines a standard hydrophobicity analysis with a detection of potential termini (edges, starts and ends) of transmembrane regions. This allows both to discard highly hydrophobic regions not delimited by clear start and end configurations and to confirm putative transmembrane segments not distinguishable by their hydrophobic composition. The accuracy obtained on a test set of 101 non homologous transmembranes proteins with reliable topologies compares well with that of other popular existing methods. Only a slight decrease in prediction accuracy was observed when the algorithm was applied to all transmembrane proteins of the SwissProt database (release 35).
Proper citation: PRED-TMR (RRID:SCR_006203) Copy
The EBI genomes pages give access to a large number of complete genomes including bacteria, archaea, viruses, phages, plasmids, viroids and eukaryotes. Methods using whole genome shotgun data are used to gain a large amount of genome coverage for an organism. WGS data for a growing number of organisms are being submitted to DDBJ/EMBL/GenBank. Genome entries have been listed in their appropriate category which may be browsed using the website navigation tool bar on the left. While organelles are all listed in a separate category, any from Eukaryota with chromosome entries are also listed in the Eukaryota page. Within each page, entries are grouped and sorted at the species level with links to the taxonomy page for that species separating each group. Within each species, entries whose source organism has been categorized further are grouped and numbered accordingly. Links are made to: * taxonomy * complete EMBL flatfile * CON files * lists of CON segments * Project * Proteomes pages * FASTA file of Proteins * list of Proteins
Proper citation: EBI Genomes (RRID:SCR_002426) Copy
http://bioinformatics.udel.edu/Research/skategenomeproject
Core facility provides a model for collaborative approaches to use specialized resources and expertise in an integrated process. Core builds on the expertise and resources provided by the Bioinformatics Cores of the five northeastern states that form NECC. The Skate Genome Annotation Workshops and Jamborees offer training and opportunities for faculty and students to work with and annotate genome sequences. Workshops include lectures, tutorials and exercises annotating the genome of the little skate, Leucoraja erinacea.
Proper citation: University of Delaware Skate Genome Project (RRID:SCR_005300) Copy
http://www.broad.mit.edu/annotation/fungi/fgi/
Produces and analyzes sequence data from fungal organisms that are important to medicine, agriculture and industry. The FGI is a partnership between the Broad Institute and the wider fungal research community, with the selection of target genomes governed by a steering committee of fungal scientists. Organisms are selected for sequencing as part of a cohesive strategy that considers the value of data from each organism, given their role in basic research, health, agriculture and industry, as well as their value in comparative genomics.
Proper citation: Fungal Genome Initiative (RRID:SCR_003169) Copy
Can't find your Tool?
We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.
Welcome to the SPARC SAWG Resources search. From here you can search through a compilation of resources used by SPARC SAWG and see how data is organized within our community.
You are currently on the Community Resources tab looking through categories and sources that SPARC SAWG has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.
If you have an account on SPARC SAWG then you can log in from here to get additional features in SPARC SAWG such as Collections, Saved Searches, and managing Resources.
Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:
You can save any searches you perform for quick access to later from here.
We recognized your search term and included synonyms and inferred terms along side your term to help get the data you are looking for.
If you are logged into SPARC SAWG you can add data records to your collections to create custom spreadsheets across multiple sources of data.
Here are the sources that were queried against in your search that you can investigate further.
Here are the categories present within SPARC SAWG that you can filter your data on
Here are the subcategories present within this category that you can filter your data on
If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.