Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
Nucleotide-binding leucine-rich-repeat (NLR) genes comprise the largest family of plant disease-resistance genes. Angiosperm NLR genes are phylogenetically divided into the TNL, CNL, and RNL subclasses. NLR copy numbers and subclass composition vary tremendously across angiosperm genomes. However, the evolutionary associations between genomic NLR content and ecological adaptation, or between NLR content and signal transduction components, are poorly characterized because of limited genome availability. In this study, we established an angiosperm NLR atlas (ANNA, https://biobigdata.nju.edu.cn/ANNA/) that includes NLR genes from over 300 angiosperm genomes. Using ANNA, we revealed that NLR copy numbers differ up to 66-fold among closely related species owing to rapid gene loss and gain. Interestingly, NLR contraction was associated with adaptations to aquatic, parasitic, and carnivorous lifestyles. The convergent NLR reduction in aquatic plants resembles the lack of NLR expansion during the long-term evolution of green algae before the colonization of land. A co-evolutionary pattern between NLR subclasses and plant immune pathway components was also identified, suggesting that immune pathway deficiencies may drive TNL loss. Finally, we identified a conserved TNL lineage that may function independently of the EDS1-SAG101-NRG1 module. Collectively, these findings provide new insights into the evolution of NLR genes in the context of ecological adaptation and genome content variation.
Pubmed ID: 34364002
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute).
View all literature mentionsWeb sevice of ClustalW provided by DNA data bank of Japan.
View all literature mentionsWeb server for whole genome comparison and annotation of orthologous clusters across multiple species.Works on any operating system with modern browser and Javascript enabled. Used to identify orthologous gene clusters and supports user define species to upload customized protein sequences. Interactive graphic tool which provides Venn diagram view for comparing multiple species protein sequences.
View all literature mentions