Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
Rhamnogalacturonan-II (RG-II) is a complex plant cell wall polysaccharide that is composed of an α(1,4)-linked homogalacturonan backbone substituted with four side chains. It exists in the cell wall in the form of a dimer that is cross-linked by a borate di-ester. Despite its highly complex structure, RG-II is evolutionarily conserved in the plant kingdom suggesting that this polymer has fundamental functions in the primary wall organisation. In this study, we have set up a bioinformatics strategy aimed at identifying putative glycosyltransferases (GTs) involved in RG-II biosynthesis. This strategy is based on the selection of candidate genes encoding type II membrane proteins that are tightly coexpressed in both rice and Arabidopsis with previously characterised genes encoding enzymes involved in the synthesis of RG-II and exhibiting an up-regulation upon isoxaben treatment. This study results in the final selection of 26 putative Arabidopsis GTs, including 10 sequences already classified in the CAZy database. Among these CAZy sequences, the screening protocol allowed the selection of α-galacturonosyltransferases involved in the synthesis of α4-GalA oligogalacturonides present in both homogalacturonans and RG-II, and two sialyltransferase-like sequences previously proposed to be involved in the transfer of Kdo and/or Dha on the pectic backbone of RG-II. In addition, 16 non-CAZy GT sequences were retrieved in the present study. Four of them exhibited a GT-A fold. The remaining sequences harbored a GT-B like fold and a fucosyltransferase signature. Based on homologies with glycosyltransferases of known functions, putative roles in the RG-II biosynthesis are proposed for some GT candidates.
Pubmed ID: 23272088
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Software platform for complex network analysis and visualization. Used for visualization of molecular interaction networks and biological pathways and integrating these networks with annotations, gene expression profiles and other state data.
View all literature mentionsThe Center for Biological Sequence Analysis of the Technical University of Denmark conducts basic research in the field of bioinformatics and systems biology and directs its research primarily towards topics related to the elucidation of the functional aspects of complex biological mechanisms. A large number of computational methods have been produced, which are offered to others via WWW servers. Several data sets are also available. The center also has experimental efforts in gene expression analysis using DNA chips and data generation in relation to the physical and structural properties of DNA. The on-line prediction services at CBS are available as interactive input forms. Most of the servers are also available as stand-alone software packages with the same functionality. In addition, for some servers, programmatic access is provided in the form of SOAP-based Web Services. The center also educates engineering students in biotechnology and systems biology and offers a wide range of courses in bioinformatics, systems biology, human health, microbiology and nutrigenomics.
View all literature mentionsDatabase of genetic and molecular biology data for the model higher plant Arabidopsis thaliana. Data available includes the complete genome sequence along with gene structure, gene product information, metabolism, gene expression, DNA and seed stocks, genome maps, genetic and physical markers, publications, and information about the Arabidopsis research community. Gene product function data is updated every two weeks from the latest published research literature and community data submissions. Gene structures are updated 1-2 times per year using computational and manual methods as well as community submissions of new and updated genes. TAIR also provides extensive linkouts from data pages to other Arabidopsis resources. The data can be searched, viewed and analyzed. Datasets can also be downloaded. Pages on news, job postings, conference announcements, Arabidopsis lab protocols, and useful links are provided.
View all literature mentionsA portal to biomedical and genomic information. NCBI creates public databases, conducts research in computational biology, develops software tools for analyzing genome data, and disseminates biomedical information for the better understanding of molecular processes affecting human health and disease.
View all literature mentionsDatabase that provides the genome sequence assembly of the International Rice Genome Sequencing Project (IRGSP), manually curated annotation of the sequence, and other genomics information that could be useful for comprehensive understanding of the rice biology. RAP-DB contains clone positions, structures and functions of genes validated by cDNAs, RNA genes detected by massively parallel signature sequencing (MPSS) technology and sequence similarity, flanking sequences of mutant lines, transposable elements, etc. Other annotation data such as Gnomon can be displayed along with those of RAP for comparison.
View all literature mentionsNASCArrays is the Nottingham Arabidopsis Stock Centre''s microarray database. Currently most of the data is for Arabidopsis thaliana experiments run by the NASC Affymetrix Facility. There are also experiments from other species, and experiments run by other centres too. NASCArrays is an Affymetrix microarray database. It contains free Affymetrix microarray data, and also features a series of tools allowing you to query that data in powerful ways. Most of the data currently comes from NASC''s Affymetrix Service. It also includes data from other sources, notably the AtGenExpress project. They currently distribute over 30,000 tubes of seed a year. There are currently the following data mining tools available. All of these tools allow you to type in a gene(s) of interest, and identify experiments or slides that you might be interested in: -Spot History: This tool allows you to see the pattern of gene expression over all slides in the database. Easily identify slides (and therefore experimental treatments) where genes are highly, lowly, or unusually expressed -Two gene scatter plot: This tool allows you to see the pattern of gene expression over all slides for two genes as a scatter plot. If you are interested in two genes, you can find out if they act in tandem, and highlight slides (and therefore experimental conditions) where these two genes behave in an unusual manner. -Gene Swinger: If you have a gene of interest, this tool will show you which experiment the gene expression varied most -Bulk Gene Download: This tool allows you to download the expression of a list of genes over all experiments. You can get all genes over all experiments (the entire database!) from the Super Bulk Gene Download Sponsors: This is a BBSRC funded consortium to provide services to the Arabidopsis community.
View all literature mentionsDatabase that describes the families of structurally-related catalytic and carbohydrate-binding modules (or functional domains) of enzymes that degrade, modify, or create glycosidic bonds. This specialist database is dedicated to the display and analysis of genomic, structural and biochemical information on Carbohydrate-Active Enzymes (CAZymes). CAZy data are accessible either by browsing sequence-based families or by browsing the content of genomes in carbohydrate-active enzymes. New genomes are added regularly shortly after they appear in the daily releases of GenBank. New families are created based on published evidence for the activity of at least one member of the family and all families are regularly updated, both in content and in description. An original aspect of the CAZy database is its attempt to cover all carbohydrate-active enzymes across organisms and across subfields of glycosciences. One can search for CAZY Family pages using the Protein Accession (Genpept Accession, Uniprot Accession or PDB ID), Cazy family name or EC number. In addition, genomes can be searched using the NCBI TaxID. This search can be complemented by Google-based searches on the CAZy site.
View all literature mentionsA database of protein families, each represented by multiple sequence alignments and hidden Markov models (HMMs). Users can analyze protein sequences for Pfam matches, view Pfam family annotation and alignments, see groups of related families, look at the domain organization of a protein sequence, find the domains on a PDB structure, and query Pfam by keywords. There are two components to Pfam: Pfam-A and Pfam-B. Pfam-A entries are high quality, manually curated families that may automatically generate a supplement using the ADDA database. These automatically generated entries are called Pfam-B. Although of lower quality, Pfam-B families can be useful for identifying functionally conserved regions when no Pfam-A entries are found. Pfam also generates higher-level groupings of related families, known as clans (collections of Pfam-A entries which are related by similarity of sequence, structure or profile-HMM).
View all literature mentionsInternational functional genomics data collection generated from microarray or next-generation sequencing (NGS) platforms. Repository of functional genomics data supporting publications. Provides genes expression data for reuse to the research community where they can be queried and downloaded. Integrated with the Gene Expression Atlas and the sequence databases at the European Bioinformatics Institute. Contains a subset of curated and re-annotated Archive data which can be queried for individual gene expression under different biological conditions across experiments. Data collected to MIAME and MINSEQE standards. Data are submitted by users or are imported directly from the NCBI Gene Expression Omnibus.
View all literature mentions