Searching the RRID Resource Information Network

Our searching services are busy right now. Please try again later

  • Register
X
Forgot Password

If you have forgotten your password you can enter your email here and get a temporary password sent to your email.

X

Leaving Community

Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.

No
Yes

Preparing word cloud

×

SciCrunch Registry is a curated repository of scientific resources, with a focus on biomedical resources, including tools, databases, and core facilities - visit SciCrunch to register your resource.

Search

Type in a keyword to search

Filter by records added date
See new records

Options


Current Facets and Filters

  • Keywords:sequence (facet)

Facets


Recent searches

Snippet view Table view
Click the to add this resource to a Collection

570 Results - per page

Show More Columns | Download 570 Result(s)

Resource Name Proper Citation Abbreviations Resource Type Description Keywords Resource Relationships Related Condition Funding Defining Citation Availability Specification URL Alternate IDs Alternate URLs Old URLs Parent Organization Resource ID Synonyms Record Last Update Mentions Count
AdapterRemoval
 
Resource Report
Resource Website
500+ mentions
AdapterRemoval (RRID:SCR_011834) software application, data analysis software, sequence analysis software, software resource, data processing software Software program to remove residual adapter sequences from next generation sequencing reads. Used for cleaning of next-generation sequencing reads. AdapterRemoval v2 introduces improvements in throughput, through use of single instruction, multiple data (SIMD; SSE1 and SSE2) instructions and multi-threading support; handles datasets containing reads or read-pairs with different adapters or adapter pairs; provides simultaneous demultiplexing and adapter trimming; has ability to reconstruct adapter sequences from paired-end reads for poorly documented data sets; provides native gzip and bzip2 support. cleaning of next-generation sequencing reads, remove residual adapter sequences, adapter, sequence, residual, next generation sequencing reads, is listed by: OMICtools
is listed by: Debian
Danish National Research Foundation ;
Lundbeck Foundation Grant ;
Marie Curie International Outgoing Fellowship within the 7th European Community Framework Programme ;
Danish Council for Independent Research
PMID:22748135
PMID:26868221
DOI:10.1186/s13104-016-1900-2
Free, Available for download, Freely available OMICS_01081 https://sources.debian.org/src/adapterremoval/ http://code.google.com/p/adapterremoval/, https://github.com/slindgreen/AdapterRemoval, https://sources.debian.org/src/adapterremoval/ SCR_011834 AdapterRemoval v2 2026-07-28 09:42:58 675
Princeton High Throughput Sequencing and Microarray Facility
 
Resource Report
Resource Website
Princeton High Throughput Sequencing and Microarray Facility (RRID:SCR_012619) Princeton High Throughput Sequencing and Microarray Facility, High Throughput Sequencing and Microarray Facility service resource, data analysis service, production service resource, access service resource, analysis service resource, core facility Core facility provides researchers with access to high-throughput sequencing technologies. The staff provide consultation on experimental design, library preparation, and data analysis. The Sequencing Core Facility works closely with Bioinformatics staff in the Center for Quantitative Biology to provide researchers with computing power and consulting services to analyze sequencing data. sequence, microarray, data analysis, analysis, consulting, is listed by: ScienceExchange
is related to: Princeton University Labs and Facilities
has parent organization: Princeton University; New Jersey; USA
Available to External User SciEx_567 SCR_012619 High Throughput Sequencing, Microarray, Princeton University, Facility 2026-07-28 09:43:13 0
Human Genome Project Information
 
Resource Report
Resource Website
50+ mentions
Human Genome Project Information (RRID:SCR_013028) narrative resource, data or information resource, slide, video resource, funding resource, topical portal, portal, training material This resource gives information about the U.S. Human Genome Project, which was was a 13-year effort to to discover all the estimated 20,000-25,000 human genes and make them accessible for further biological study. The primary project goals were to: - identify all the approximately 20,000-25,000 genes in human DNA, - determine the sequences of the 3 billion chemical base pairs that make up human DNA, - store this information in databases, - improve tools for data analysis, - transfer related technologies to the private sector, and - address the ethical, legal, and social issues (ELSI) that may arise from the project. To help achieve these goals, researchers also studied the genetic makeup of several nonhuman organisms. These include the common human gut bacterium Escherichia coli, the fruit fly, and the laboratory mouse. These parallel studies helped to develop technology and interpret human gene function. Sponsors: The DOE Human Genome Program and the NIH National Human Genome Research Institute (NHGRI) together sponsored the U.S. Human Genome Project. escherichia coli, fruit fly, function, gene, genome, genetic, bacterium, base pair, biological, dna, human, mouse, sequence, FASEB list has parent organization: National Institutes of Health
has parent organization: United States Department of Energy
nif-0000-10252 SCR_013028 HGP 2026-07-28 09:43:21 59
SIFT
 
Resource Report
Resource Website
10000+ mentions
SIFT (RRID:SCR_012813) SIFT service resource, data analysis service, data access protocol, software resource, source code, production service resource, web service, analysis service resource Data analysis service to predict whether an amino acid substitution affects protein function based on sequence homology and the physical properties of amino acids. SIFT can be applied to naturally occurring nonsynonymous polymorphisms and laboratory-induced missense mutations. (entry from Genetic Analysis Software) Web service is also available. gene, genetic, genomic, amino acid, substitution, protein function, coding region, single nucleotide variant, coding indel, deletion, insertion, sequence, protein, bio.tools is listed by: OMICtools
is listed by: Genetic Analysis Software
is listed by: Debian
is listed by: bio.tools
is listed by: SoftCite
is related to: SIFT 4G
has parent organization: Genome Institute of Singapore; Singapore; Singapore
has parent organization: J. Craig Venter Institute
Agency for Science Technology and Research ;
NIGMS GM29009
PMID:19561590
PMID:12824425
PMID:11337480
DOI:10.1038/nprot.2009.86
Non-commercial biotools:sift, OMICS_00137, nlx_154618 http://sift.jcvi.org/, https://bio.tools/sift, https://sources.debian.org/src/sift/ http://sift.bii.a-star.edu.sg/SIFT.html SCR_012813 Sorting Intolerant From Tolerant 2026-07-28 09:43:10 10223
DOLOP: A Database of Bacterial Lipoproteins
 
Resource Report
Resource Website
10+ mentions
DOLOP: A Database of Bacterial Lipoproteins (RRID:SCR_013487) service resource, data or information resource, data repository, database, storage service resource DOLOP is an exclusive knowledge base for bacterial lipoproteins by processing information from 510 entries to provide a list of 199 distinct lipoproteins with relevant links to molecular details. Features include functional classification, predictive algorithm for query sequences, primary sequence analysis and lists of predicted lipoproteins from 43 completed bacterial genomes along with interactive information exchange facility. This website along will have additional information on the biosynthetic pathway, supplementary material and other related figures. DOLOP also contains information and links to molecular details for about 278 distinct lipoproteins and predicted lipoproteins from 234 completely sequenced bacterial genomes. Additionally, the website features a tool that applies a predictive algorithm to identify the presence or absence of the lipoprotein signal sequence in a user-given sequence. The experimentally verified lipoproteins have been classified into different functional classes and more importantly functional domain assignments using hidden Markov models from the SUPERFAMILY database that have been provided for the predicted lipoproteins. Other features include: primary sequence analysis, signal sequence analysis, and search facility and information exchange facility to allow researchers to exchange results on newly characterized lipoproteins. figure, functional, algorithm, analysis, bacterial, biosynthetic, classification, genome, lipid, lipoprotein, modification, molecular, molecule, pathogenesis, predictive, primary, prokaryote, query, sequence, signal has parent organization: University of Cambridge; Cambridge; United Kingdom nif-0000-21124 SCR_013487 DOLOP 2026-07-28 09:43:17 16
PolyPhen: Polymorphism Phenotyping
 
Resource Report
Resource Website
1000+ mentions
PolyPhen: Polymorphism Phenotyping (RRID:SCR_013189) PolyPhen, PolyPhen-2, POLYPHEN software application, data analysis software, software resource, data processing software, simulation software Software tool which predicts possible impact of amino acid substitution on structure and function of human protein using straightforward physical and comparative considerations. PolyPhen-2 is new development of PolyPhen tool for annotating coding nonsynonymous SNPs. annotate, nonsynonymous, SNP, predict, coding, damaging, effect, missense, mutation, sequence, variant, phenotype, genetic, disease, exon, protein, coding, fraction, genome, bio.tools is listed by: Genetic Analysis Software
is listed by: Debian
is listed by: bio.tools
is related to: OMICtools
has parent organization: Harvard University; Cambridge; United States
PMID:20354512
PMID:23315928
SCR_013200, OMICS_00136, nlx_154540, nif-0000-21329, biotools:polyphen, SCR_013238 https://bio.tools/polyphen http://www.bork.embl-heidelberg.de/PolyPhen/ SCR_013189 PolyPhen, POLYPHEN, PolyPhen-2, Polymorphism Phenotyping, Polymorphism Phenotyping v2 2026-07-28 09:43:11 4151
SeqEM
 
Resource Report
Resource Website
1+ mentions
SeqEM (RRID:SCR_002021) software application, web application, data analysis software, algorithm resource, sequence analysis software, software resource, data processing software Online tool for utilizing a genotype calling algorithm for next-generation sequence data. genotype, algorithm, sequence, rna, dna, bio.tools is listed by: OMICtools
is listed by: bio.tools
is listed by: Debian
has parent organization: University of Miami Miller School of Medicine; Florida; USA
PMID:20861027 THIS RESOURCE IS NO LONGER IN SERVICE OMICS_00074, biotools:seqem https://bio.tools/seqem SCR_002021 2026-07-28 09:40:19 1
RNA Ontology
 
Resource Report
Resource Website
1+ mentions
RNA Ontology (RRID:SCR_003470) RNAO controlled vocabulary, data or information resource, ontology An ontology to capture all aspects of RNA - from primary sequence to alignments, secondary and tertiary structure from base pairing and base stacking to sophisticated motifs. owl, obo, molecular structure, molecular, rna, sequence, alignment, structure, base pairing, base stacking, motif is listed by: BioPortal
is listed by: OBO
is listed by: Google Code
Free, Available for download, Freely available nlx_157566 http://purl.bioontology.org/ontology/RNAO, http://rnao.googlecode.com/svn/trunk/rnao.obo SCR_003470 2026-07-28 09:40:48 1
Influenza Ontology
 
Resource Report
Resource Website
Influenza Ontology (RRID:SCR_003346) FLU controlled vocabulary, data or information resource, ontology An application ontology established by a collaborative group of influenza researchers that includes consolidated influenza sequence and surveillance terms from resources such as the BioHealthBase (BHB), a Bioinformatics Resource Center (BRC) for Biodefense and Emerging and Re-emerging Infectious Diseases, the Centers for Excellence in Influenza Research and Surveillance (CEIRS) owl, health, pathological, organismal, cellular, sequence, surveillance is listed by: BioPortal
is listed by: OBO
is related to: Information Artifact Ontology
has parent organization: University of Maryland; Maryland; USA
Influenza Free, Freely available nlx_157440 http://purl.obolibrary.org/obo/flu.owl, http://influenzaontologywiki.igs.umaryland.edu/ http://purl.bioontology.org/ontology/FLU SCR_003346 2026-07-28 09:40:40 0
COnsensus-DEgenerate Hybride Oligonucleotide Primers
 
Resource Report
Resource Website
1+ mentions
COnsensus-DEgenerate Hybride Oligonucleotide Primers (RRID:SCR_002875) software application, service resource, data analysis software, data analysis service, data processing software, software resource, production service resource, analysis service resource This COnsensus-DEgenerate Hybrid Oligonucleotide Primer (CODEHOP) strategy has been implemented as a computer program that is accessible over the World-Wide Web and is directly linked from the BlockMaker multiple sequence alignment site for hybrid primer prediction beginning with a set of related protein sequences. This is a new primer design strategy for PCR amplification of unknown targets that are related to multiply-aligned protein sequences. Each primer consists of a short 3' degenerate core region and a longer 5' consensus clamp region. Only 3-4 highly conserved amino acid residues are necessary for design of the core, which is stabilized by the clamp during annealing to template molecules. During later rounds of amplification, the non-degenerate clamp permits stable annealing to product molecules. The researchers demonstrate the practical utility of this hybrid primer method by detection of diverse reverse transcriptase-like genes in a human genome, and by detection of C5 DNA methyltransferase homologs in various plant DNAs. In each case, amplified products were sufficiently pure to be cloned without gel fractionation. Sponsors: This work was supported in part by a grant from the M. J. Murdock Charitable Trust and by a grant from NIH. S. P. is a Howard Hughes Medical Institute Fellow of the Life Sciences Research Foundation., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on January 15,2026. fractionation, gel, 3', amplification, clone, dna, genome, homolog, human, hybrid, molecule, oligonucleotide, pcr, plant, primer, protein, sequence, transcriptase-methyltransferase is related to: OMICtools
has parent organization: University of Washington; Seattle; USA
THIS RESOURCE IS NO LONGER IN SERVICE nif-0000-25557 SCR_002875 CODEHOP 2026-07-28 09:40:32 8
LAST
 
Resource Report
Resource Website
100+ mentions
LAST (RRID:SCR_006119) LAST software application, service resource, data analysis service, software resource, data processing software, production service resource, analysis service resource THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software tool for aligning sequences, similar to BLAST 2 sequences that colour-codes the alignments by reliability. Another useful feature of LAST is that it can compare huge (vertebrate-genome-sized) datasets. Unfortunately, this only applies to the downloadable version of LAST, not the web service. The web service can just about handle bacterial genomes, but it will take a few minutes and the output will be large. LAST can: * Handle big sequence data, e.g: ** Compare two vertebrate genomes ** Align billions of DNA reads to a genome * Indicate the reliability of each aligned column. * Use sequence quality data properly. * Compare DNA to proteins, with frameshifts. * Compare PSSMs to sequences * Calculate the likelihood of chance similarities between random sequences. LAST cannot (yet): * Do spliced alignment., THIS RESOURCE IS NO LONGER IN SERVICE. Documented on September 16,2025. sequence alignment, align, vertebrate, genome, sequence, alignment, bio.tools is listed by: OMICtools
is listed by: bio.tools
is listed by: Debian
is related to: RecountDB
has parent organization: National Institute of Advanced Industrial Science and Technology
National Genome Research Network ;
INTEuropean Union Systems Institute ;
Japanese Ministry of Education Culture Sports Science and Technology MEXT
PMID:21209072
PMID:20144198
PMID:20110255
DOI:10.1093/nar/gkq010
THIS RESOURCE IS NO LONGER IN SERVICE nlx_151594, biotools:last, OMICS_15813 https://bio.tools/last, https://sources.debian.org/src/last-align/ SCR_006119 2026-07-28 09:41:31 397
Protein Classification Benchmark Collection
 
Resource Report
Resource Website
10+ mentions
Protein Classification Benchmark Collection (RRID:SCR_007561) data or information resource, database It was created in order to create standard datasets on which the performance of machine learning methods can be compared. The collection contains datasets of sequences and structures, each subdivided into positive/negative training/test sets. Such a subdivision is called a classification task. Typical tasks include the classification of structural domains in the SCOP and CATH databases based on their sequences, as fell as various functional and taxonomic classification tasks. Running a performance evaluation test on an entire database can include many different classification tasks. These ensembles of classification tasks are encoded in a simple matrix format - called the cast matrix or membership table - that specifies the role of each sequence (or structure) in the different calculations. Each column of this matrix is a subdivision of the objects (rows) into positive/negative training/test sets. Typically, a database record contains such an ensemble of classification tasks, encoded in a single cast matrix. In addition, there is a collection of distance matrices that contain an all vs. all comparison of the datasets using methods as BLAST, Smith-Waterman, 3D-comparisons etc. Evaluation of a method on a given database consists of calculating a performance measure such as a receiver operating curve (ROC) AUC value. Results of evaluation are deposited along with the data, each dataset is evaluated at least by one classification method, such as 1NN (nearest neighbour) or SVM (support vector machines), ANN (artificial neural networks), RF (random forests) etc.. There are small datasets meant for program developers, as well as downloadable programs for various classification algorithms. classificaiton, machine learning, sequence, standard, standard dataset, structure, technology nif-0000-02600 SCR_007561 Benchmark 2026-07-28 09:42:05 24
CMKB
 
Resource Report
Resource Website
1+ mentions
CMKB (RRID:SCR_007229) CMKB data or information resource, database It is a database of keys facts about proteins, families, and complexes involved in cell migration. This ongoing project provides a large amount of automated and curated data, collected from numerous online resources that are updated monthly. These data include names, synonyms, sequence information, summaries, CMC research data, reagents, structures, as well as protein family and complex details. CMKB''s ultimate goal is to create a database that will enable the cell migration community to conveniently access significant information about molecules of interest. This will also serve as a stepping stone to pathway analysis and demonstrate how these molecules coordinate with one another during cell adhesion and movement. Sponsors: This resource is supported by the Cell Migration Consortium. cell, migration, knowledgebase, database, protein, family, data, sequence, synonyms, research, reagent, structure, protein, molecule, interest, pathway, adhesion, movement nif-0000-30312 SCR_007229 Cell Migration Knowledgebase, The Cell Migration Knowledgebase 2026-07-28 09:41:55 1
Alternate splicing gallery
 
Resource Report
Resource Website
1+ mentions
Alternate splicing gallery (RRID:SCR_008129) data or information resource, database Alternative splicing essentially increases the diversity of the transcriptome and has important implications for physiology, development and the genesis of diseases. This resource uses a different approach to investigate alternative splicing (instead of the conventional case-by case fashion) and integrates all transcripts derived from a gene into a single splicing graph. ASG is a database of splicing graphs for human genes, using transcript information from various major sources (Ensembl, RefSeq, STACK, TIGR and UniGene). Each transcript corresponds to a path in the graph, and alternative splicing is displayed by bifurcations. This representation preserves the relationships between different splicing variants and allows us to investigate systematically all possible putative transcripts. Web interface allows users to display the splicing graphs, to interactively assemble transcripts and to access their sequences as well as neighboring genomic regions. ASG also provide for each gene, an exhaustive pre-computed catalog of putative transcriptsin total more than 1.2 million sequences. It has found that ~65 of the investigated genes show evidence for alternative splicing, and in 5 of the cases, a single gene might produce over 100 transcripts. gallery, gene, genesis, alternative, development, disease, diversity, genomic, human, physiology, putative transcript, sequence, single, splice, splicing graph, transcript, transcriptome, variant, bio.tools is listed by: bio.tools
is listed by: Debian
nif-0000-20932, biotools:alternative_splicing_gallery https://bio.tools/alternative_splicing_gallery SCR_008129 ASG 2026-07-28 09:42:03 1
Kinase Pathway Database
 
Resource Report
Resource Website
1+ mentions
Kinase Pathway Database (RRID:SCR_008199) data or information resource, database THIS RESOURCE IS NO LONGER IN SERVICE, documented August 23, 2016. KinasePathwayDatabase is an integrated database concerning completed sequenced major eukaryotes, which contains the classification of protein kinases and their functional conservation and orthologous tables among species, protein-protein interaction data, domain information, structural information, and automatic pathway graph image interface. The protein-protein interactions are extracted by natural language processing (NLP) from abstracts using basic word pattern and protein name dictionary GENA: developed by our group. In this system, pathways are easily compared among species using protein interactions data more than 47,000 and orthologous tables. eukaryote, functional, automatic, classification, conservation, domain, interaction, intermolecular interactions and signaling pathways databases, kinase, orthologous, pathway, protein, sequence, specie, structural, image has parent organization: University of Tokyo; Tokyo; Japan THIS RESOURCE IS NO LONGER IN SERVICE nif-0000-21235 SCR_008199 Kinase Pathway Database 2026-07-28 09:42:17 2
The Protein Coil Library
 
Resource Report
Resource Website
1+ mentions
The Protein Coil Library (RRID:SCR_008233) data or information resource, database The Protein Coil Library is a library of protein structure fragments derived from the Protein Data Bank (PDB). The fragments in this library are those fragments in the PDB that cannot be classified as either alpha-helix or beta-strand. Three-dimensional structures as well as side-chain and backbone torsion angles are stored in the database. The Protein Coil Library allows rapid and comprehensive access to non-alpha-helix and non-beta-strand fragments contained in the Protein Data Bank (PDB). The library contains both sequence and structure information together with calculated torsion angles for both the backbone and side chains. Several search options are implemented, including a query function that uses output from popular PDB-culling servers directly. Additionally, several popular searches are stored and updated for immediate access. The library is a useful tool for exploring conformational propensities, turn motifs, and a recent model of the unfolded state. The library stores the complete torsion angle descriptions for the fragments as well as the three dimensional structures of the fragments themselves. The goal of extracting and pre-calculating this data is to allow for more straightforward investigation of peptide structure without the background of secondary structure elements. In addition to searching by PDB ID, it is possible to download a particular size class, perform a batch search of PDB/chain ID''s, or download precompiled lists of PDB ID''s of interest (PDB Select, etc.). For users interested in browsing the entire database at once or maintaining their own locally-updated copy of the library, FTP access instructions are also provided. The files stored in the coil library FTP site or returned after a batch search are organized heirarchically by PDB ID. This is done to reduce filesystem access times and fascilitate searches using the UNIX find utility. At the lowest directory level in the heirarchy, files are further sorted by fragment length. As a result, the number of files in a particular directory is generally less then 50, yielding relatively fast access on UNIX/Linux filesystems. The heirarchical organization is based on the middle two letters of the PDB ID. For example, hen egg lysozyme, which has a PDB ID of 1HEL, will be located in the directory h/he/. At the final level, fragments of varying sizes are stored in directories that correspond to their fragment length. Again, using lysozyme as an example, any seven-residue fragments, if they exist, will reside in the directory h/he/7/. Similarly, seven-residue fragments from 2HEX and 1HE0 will also be in this location. Sponsors: The Protein Coil Library is funded by Johns Hopkins University. element, fragment, alpha helix, angle, backbone, beta, coil, lysozyme, peptide, protein, protein structure databases, secondary, sequence, side chain, strand, structure, torsion has parent organization: Johns Hopkins University; Maryland; USA nif-0000-21338 SCR_008233 The Coil Library 2026-07-28 09:42:03 1
Ancient conserved untranslated sequences
 
Resource Report
Resource Website
Ancient conserved untranslated sequences (RRID:SCR_008130) ACUTS data or information resource, database THIS RESOURCE IS NO LONGER IN SERVICE, Documented on August 12, 2014. Database that identifies new regulatory elements in untranslated regions of protein-coding genes (5 prime flanks, 5 prime UTRs, introns, 3 prime UTRs and 3 prime flanks). The analyses is focused on genes from metazoan species (essentially vertebrates, insects and nematodes). Information on highly conserved regions (sequences, alignments, annotations, bibliographic references) are compiled. Currently 176 out of 326 detected highly conserved regions (HCRs) have been analyzed and incorporated in the database. You can also access the list of annotated conserved elements and the list of conserved elements that remain to be processed. Their approach is based on comparative sequence analysis, for the identification of phylogenetic footprints. echinoderm, footprint, fragment, functional, gene, alignment, analysis, annotation, chordate, cis-element, coding, degradation, divergence, dna, dnase, highly conserved region, homologous, intron, metazoan, mrna, non-coding, nucleotide, phylogenetic, post-transcriptional, promoter, protein, region, regulatory, segment, sequence, structural, transcriptional repressor, translation, untranslated region has parent organization: Claude Bernard University Lyon 1; Lyon; France PMID:9204283 THIS RESOURCE IS NO LONGER IN SERVICE nif-0000-20934 SCR_008130 2026-07-28 09:42:17 0
Animal Genome Database
 
Resource Report
Resource Website
1+ mentions
Animal Genome Database (RRID:SCR_008165) data or information resource, database Database of comparative gene mapping between species to assist the mapping of the genes related to phenotypic traits in livestock. The linkage maps, cytogenetic maps, polymerase chain reaction primers of pig, cattle, mouse and human, and their references have been included in the database, and the correspondence among species have been stipulated in the database. AGP is an animal genome database developed on a Unix workstation and maintained by a relational database management system. It is a joint project of National Institute of Agrobiological Sciences (NIAS) and Institute of the Society for Techno-innovation of Agriculture, Forestry and Fisheries (STAFF-Institute), under cooperation with other related research institutes. AGP also contains the Pig Expression Data Explorer (PEDE), a database of porcine EST collections derived from full-length cDNA libraries and full-length sequences of the cDNA clones picked from the EST collection. The EST sequences have been clustered and assembled, and their similarity to sequences in RefSeq, and UniGene determined. The PEDE database system was constructed to store sequences and similarity data of swine full-length cDNA libraries and to make them available to users. It provides interfaces for keyword and ID searches of BLAST results and enables users to obtain sequence data and names of clones of interest. Putative SNPs in EST assemblies have been classified according to breed specificity and their effect on coding amino acids, and the assemblies are equipped with an SNP search interface. The database contains porcine nucleotide sequences and cDNA clones that are ready for analyses such as expression in mammalian cells, because of their high likelihood of containing full-length CDS. PEDE will be useful for researchers who want to explore genes that may be responsible for traits such as disease susceptibility. The database also offers information regarding major and minor porcine-specific antigens, which might be investigated in regard to the use of pigs as models in various medical research applications. est, expression, gene, amino acid, animal, antigen, breed, cattle, cdna, cell, chain, clone, coding, cytogenetic, genome, human, linkage, livestock, mammalian, map, mouse, nucleotide, organism, phenotypic, pig, polymerase, porcine, primer, reaction, sequence, snp, specie, swine, trait has parent organization: National Institute of Agrobiological Sciences; Ibaraki; Japan nif-0000-21029 SCR_008165 AGP 2026-07-28 09:42:17 1
Phylogenetic Clusters of Orthologous Groups Ranking
 
Resource Report
Resource Website
1+ mentions
Phylogenetic Clusters of Orthologous Groups Ranking (RRID:SCR_008223) data or information resource, database THIS RESOURCE IS NO LONGER IN SERVICE, documented on August 20,2019.The COG-database has become a powerful tool in the field of comparative genomics. The construction of this data-base is based on sequence homologies of proteins from different completely sequenced genomes. Highly homologous proteins are assigned to clusters of orthologous groups. The updated collection of orthologous protein sets for prokaryotes and eukaryotes is expected to be a useful platform for functional annotation of newly sequenced genomes, including those of complex eukaryotes, and genome-wide evolutionary studies. The availability of multiple, essentially complete genome sequences of prokaryotes and eukaryotes spurred both the demand and the opportunity for the construction of an evolutionary classification of genes from these genomes. Such a classification system based on orthologous relationships between genes appears to be a natural framework for comparative genomics and should facilitate both functional annotation of genomes and large-scale evolutionary studies. Here is a major update of the previously developed system for delineation of Clusters of Orthologous Groups of proteins (COGs) from the sequenced genomes of prokaryotes and unicellular eukaryotes and the construction of clusters of predicted orthologs for 7 eukaryotic genomes, which we named KOGs after eukaryotic orthologous groups. The COG collection currently consists of 138,458 proteins, which form 4873 COGs and comprise 75% of the 185,505 (predicted) proteins encoded in 66 genomes of unicellular organisms. The eukaryotic orthologous groups (KOGs) include proteins from 7 eukaryotic genomes: three animals (the nematode Caenorhabditis elegans, the fruit fly Drosophila melanogaster and Homo sapiens), one plant, Arabidopsis thaliana, two fungi (Saccharomyces cerevisiae and Schizosaccharomyces pombe), and the intracellular microsporidian parasite Encephalitozoon cuniculi. The current KOG set consists of 4852 clusters of orthologs, which include 59,838 proteins, or approximately 54% of the analyzed eukaryotic 110,655 gene products. Compared to the coverage of the prokaryotic genomes with COGs, a considerably smaller fraction of eukaryotic genes could be included into the KOGs; addition of new eukaryotic genomes is expected to result in substantial increase in the coverage of eukaryotic genomes with KOGs. Examination of the phyletic patterns of KOGs reveals a conserved core represented in all analyzed species and consisting of approximately 20% of the KOG set. This conserved portion of the KOG set is much greater than the ubiquitous portion of the COG set (approximately 1% of the COGs). In part, this difference is probably due to the small number of included eukaryotic genomes, but it could also reflect the relative compactness of eukaryotes as a clade and the greater evolutionary stability of eukaryotic genomes. elegans, encephalitozoon, eukaryote, evolutionary, fly, fruit, fungus, gene, general genomics databases, animal, arabidopsis, caenorhabditis, cerevisiae, classification, comparative, cuniculi, drosophila, genome, genomic, homo, homology, intracellular, melanogaster, microsporidian, nematode, organism, ortholog, orthologous, parasite, pattern, phyletic, phylogenetic, plant, pombe, prokaryote, protein, saccharomyces, sapiens, schizosaccharomyces, sequence, thaliana, tool, unicellular has parent organization: National Institutes of Health THIS RESOURCE IS NO LONGER IN SERVICE nif-0000-21313 SCR_008223 PCOGR 2026-07-28 09:42:18 3
Comparative Vertebrate Sequencing
 
Resource Report
Resource Website
Comparative Vertebrate Sequencing (RRID:SCR_008213) data or information resource, database Generates data for use in developing and refining computational tools for comparing genomic sequence from multiple species. The NISC Comparative Sequencing Program's goal is to establish a data resource consisting of sequences for the same set of targeted genomic regions derived from multiple animal species. The broader program includes plans for a diverse set of analytical studies using the generated sequence and the publication of a series of papers describing the results of those analysis in peer-reviewed journals in a timely fashion. Experimentally, this project involves the shotgun sequencing of mapped BAC clones. For each BAC, an assembly is first performed when a sufficient number of sequence reads have been generated to provide full shotgun coverage of the clone. At that time, the assembled sequence is submitted to the HTGS division of GenBank. Subsequent refinements of the sequence, including the generation of higher-accuracy finished sequence, results in the updating of the sequence record in GenBank. By immediately submitting our BAC-derived sequences to GenBank, it makes their data available as a public service to allow colleagues to speed up their research, consistent with the now well-established routine of sequencing centers participating in the Human Genome Project. However, at the same time, it has made considerable investment in acquiring these mapping and sequence data, including sizable efforts of graduate students, postdoctoral fellows, and other trainees. Furthermore, in most cases, large data sets involving multiple BAC sequences from multiple species must first be generated, often taking many months to accumulate, before the planned analysis can be performed and the resulting papers written and submitted for publication. accuracy, animal, bac, clone, comparative, computational, genome, genomic, human, map, mapping, model organisms and comparative genomics databases, sequence, specie, tool has parent organization: National Institutes of Health nif-0000-21291 SCR_008213 Comparative Vertebrate Sequencing 2026-07-28 09:42:02 0

Can't find your Tool?

We recommend that you click next to the search bar to check some helpful tips on searches and refine your search firstly. Alternatively, please register your tool with the SciCrunch Registry by adding a little information to a web form, logging in will enable users to create a provisional RRID, but it not required to submit.

Can't find the RRID you're searching for? X
X
  1. SPARC Anatomical Working Group Resources

    Welcome to the SPARC SAWG Resources search. From here you can search through a compilation of resources used by SPARC SAWG and see how data is organized within our community.

  2. Navigation

    You are currently on the Community Resources tab looking through categories and sources that SPARC SAWG has compiled. You can navigate through those categories from here or change to a different tab to execute your search through. Each tab gives a different perspective on data.

  3. Logging in and Registering

    If you have an account on SPARC SAWG then you can log in from here to get additional features in SPARC SAWG such as Collections, Saved Searches, and managing Resources.

  4. Searching

    Here is the search term that is being executed, you can type in anything you want to search for. Some tips to help searching:

    1. Use quotes around phrases you want to match exactly
    2. You can manually AND and OR terms to change how we search between words
    3. You can add "-" to terms to make sure no results return with that term in them (ex. Cerebellum -CA1)
    4. You can add "+" to terms to require they be in the data
    5. Using autocomplete specifies which branch of our semantics you with to search and can help refine your search
  5. Collections

    If you are logged into SPARC SAWG you can add data records to your collections to create custom spreadsheets across multiple sources of data.

  6. Facets

    Here are the facets that you can filter the data by.

  7. Further Questions

    If you have any further questions please check out our FAQs Page to ask questions and see our tutorials. Click this button to view this tutorial again.