Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
About 1 million expressed sequence tag (EST) sequences comprising 125.3 Mb nucleotides were accreted from 51 cDNA libraries constructed from a variety of tissues and organs under a range of conditions, including abiotic stresses and pathogen challenges in common wheat (Triticum aestivum). Expressed sequence tags were assembled with stringent parameters after processing with inbuild scripts, resulting in 37,138 contigs and 215,199 singlets. In the assembled sequences, 10.6% presented no matches with existing sequences in public databases. Functional characterization of wheat unigenes by gene ontology annotation, mining transcription factors, full-length cDNA, and miRNA targeting sites were carried out. A bioinformatics strategy was developed to discover single-nucleotide polymorphisms (SNPs) within our large EST resource and reported the SNPs between and within (homoeologous) cultivars. Digital gene expression was performed to find the tissue-specific gene expression, and correspondence analysis was executed to identify common and specific gene expression by selecting four biotic stress-related libraries. The assembly and associated information cater a framework for future investigation in functional genomics.
Pubmed ID: 22334568
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Collection of data of protein sequence and functional information. Resource for protein sequence and annotation data. Consortium for preservation of the UniProt databases: UniProt Knowledgebase (UniProtKB), UniProt Reference Clusters (UniRef), and UniProt Archive (UniParc), UniProt Proteomes. Collaboration between European Bioinformatics Institute (EMBL-EBI), SIB Swiss Institute of Bioinformatics and Protein Information Resource. Swiss-Prot is a curated subset of UniProtKB.
View all literature mentionsNIH genetic sequence database that provides annotated collection of all publicly available DNA sequences for almost 280 000 formally described species (Jan 2014) .These sequences are obtained primarily through submissions from individual laboratories and batch submissions from large-scale sequencing projects, including whole-genome shotgun (WGS) and environmental sampling projects. Most submissions are made using web-based BankIt or standalone Sequin programs, and GenBank staff assigns accession numbers upon data receipt. It is part of International Nucleotide Sequence Database Collaboration and daily data exchange with European Nucleotide Archive (ENA) and DNA Data Bank of Japan (DDBJ) ensures worldwide coverage. GenBank is accessible through NCBI Entrez retrieval system, which integrates data from major DNA and protein sequence databases along with taxonomy, genome, mapping, protein structure and domain information, and biomedical journal literature via PubMed. BLAST provides sequence similarity searches of GenBank and other sequence databases. Complete bimonthly releases and daily updates of GenBank database are available by FTP.
View all literature mentionsA web-based browser for Gene Ontology terms and annotations, which is provided by the UniProtKB-GOA group at the EBI. It is able to offer a range of facilities including bulk downloads of GO annotation data which can be extensively filtered by a range of different parameters and GO slim set generation. The software for QuickGO is freely available under the Apache 2 license. QuickGO can supply GO term information and GO annotation data via REST web services.
View all literature mentionsDatabase of genetic and molecular biology data for the model higher plant Arabidopsis thaliana. Data available includes the complete genome sequence along with gene structure, gene product information, metabolism, gene expression, DNA and seed stocks, genome maps, genetic and physical markers, publications, and information about the Arabidopsis research community. Gene product function data is updated every two weeks from the latest published research literature and community data submissions. Gene structures are updated 1-2 times per year using computational and manual methods as well as community submissions of new and updated genes. TAIR also provides extensive linkouts from data pages to other Arabidopsis resources. The data can be searched, viewed and analyzed. Datasets can also be downloaded. Pages on news, job postings, conference announcements, Arabidopsis lab protocols, and useful links are provided.
View all literature mentionsCollection of data related to crop plant and model organism Zea mays. Used to synthesize, display, and provide access to maize genomics and genetics data, prioritizing mutant and phenotype data and tools, structural and genetic map sets, and gene models and to provide support services to the community of maize researchers. Data stored at MaizeGDB was inherited from the MaizeDB and ZmDB projects. Sequence data are from GenBank. Data are searchable by phenotype, traits, Pests, Gel Pattern, and Mutant Images.
View all literature mentionsDatabase that provides the genome sequence assembly of the International Rice Genome Sequencing Project (IRGSP), manually curated annotation of the sequence, and other genomics information that could be useful for comprehensive understanding of the rice biology. RAP-DB contains clone positions, structures and functions of genes validated by cDNAs, RNA genes detected by massively parallel signature sequencing (MPSS) technology and sequence similarity, flanking sequences of mutant lines, transposable elements, etc. Other annotation data such as Gnomon can be displayed along with those of RAP for comparison.
View all literature mentionsDatabase and resource that provides sequence and annotation data for the rice genome. This website provides genome sequence from the Nipponbare subspecies of rice and annotation of the 12 rice chromosomes. All structural and functional annotation is viewable through our Rice Genome Browser which currently supports 75 tracks of annotation. Enhanced data access is available through web interfaces, FTP downloads and a Data Extractor tool developed in order to support discrete dataset downloads. Rice is a model species for the monocotyledonous plants and the cereals which are the greatest source of food for the world''s population. While rice genome sequence is available through multiple sequencing projects, high quality, uniform annotation is required in order for genome sequence data to be fully utilized by researchers. The existence of a common gene set and uniform annotation allows researchers within the rice community to work from a common resource so that their results can be more easily interpreted by other scientists. The objective of this project has always been to provide high quality annotation for the rice genome. They generated, refined and updated gene models for the estimated 40,000-60,000 total rice genes, provided standardized annotation for each model, linked each model to functional annotation including expression data, gene ontologies, and tagged lines. They have provided a resource to extend the annotation of the rice genome to other plant species by providing comparative alignments to other plant species. Analysis/Tools are available including: BLAST, Locus Name Search, Functional Term Search, Protein Domain Search, Anatomy Expression Viewer, Highly Expressed Genes
View all literature mentionsA computational biology laboratory that builds and redistributes genetic software tools.
View all literature mentionsA plant small RNA target analysis server which features two important analysis functions: 1) reverse complementary matching between miRNA and target transcript using a proven scoring schema, and 2) target site accessibility evaluation by calculating unpaired energy (UPE) required to ?open? secondary structure around miRNA?s target site on mRNA. PsRNATarget incorporates recent discoveries in plant miRNA target recognition, e.g. it distinguishes translational and post-transcriptional inhibition, and it reports the number of miRNA/target site pairs that may affect miRNA binding activity to target transcript. PsRNATarget is designed for high-throughput analysis of next-generation data with an efficient distributed computing back-end pipeline that runs on a Linux cluster. The server front-end integrates three simplified user-friendly interfaces to accept user-submitted or preloaded miRNAs and transcript sequences; and outputs a comprehensive list of miRNA / target pairs along with the online tools for batch downloading, key word searching and results sorting.
View all literature mentions