Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
SAR202 bacteria in the Chloroflexota phylum are abundant and widely distributed in the ocean. Their genome coding capacities indicate their potential roles in degrading complex and recalcitrant organic compounds in the ocean. However, our understanding of their genomic diversity, vertical distribution, and depth-related metabolisms is still limited by the number of assembled SAR202 genomes. In this study, we apply deep metagenomic sequencing (180 Gb per sample) to investigate microbial communities collected from six representative depths at the Bermuda Atlantic Time Series (BATS) station. We obtain 173 SAR202 metagenome-assembled genomes (MAGs). Intriguingly, 154 new species and 104 new genera are found based on these 173 SAR202 genomes. We add 12 new subgroups to the current SAR202 lineages. The vertical distribution of 20 SAR202 subgroups shows their niche partitioning in the euphotic, mesopelagic, and bathypelagic oceans, respectively. Deep-ocean SAR202 bacteria contain more genes and exhibit more metabolic potential for degrading complex organic substrates than those from the euphotic zone. With deep metagenomic sequencing, we uncover many new lineages of SAR202 bacteria and their potential functions which greatly deepen our understanding of their diversity, vertical profile, and contribution to the ocean's carbon cycling, especially in the deep ocean.
Pubmed ID: 38997445
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Quality assessment software tool for evaluating and comparing genome assemblies. It works both with and without a given reference genome. It produces many reports, summary tables and plots.
View all literature mentionsA database of orthologous groups of genes. The orthologous groups are annotated with functional description lines (derived by identifying a common denominator for the genes based on their various annotations), with functional categories (i.e derived from the original COG/KOG categories). eggNOG's database currently counts 1.7 million orthologous groups in 3686 species, covering over 7.7 million proteins (built from 9.6 million proteins). (Jan 30, 2014)
View all literature mentionsDatabases of protein sequences and 3D structures of proteins. Collection of sequences from several sources, including translations from annotated coding regions in GenBank, RefSeq and TPA, as well as records from SwissProt, PIR, PRF, and PDB.
View all literature mentionsSoftware Java pipeline for trimming tasks for Illumina paired end and single ended data. Flexible Trimmer for Illumina Sequence Data. Pair aware preprocessing tool optimized for Illumina next generation sequencing data. Includes several processing steps for read trimming and filtering. Operating systems Unix/Linux, Mac OS, Windows.
View all literature mentionsIntegrated database resource consisting of 16 main databases, broadly categorized into systems information, genomic information, and chemical information. In particular, gene catalogs in completely sequenced genomes are linked to higher-level systemic functions of cell, organism, and ecosystem. Analysis tools are also available. KEGG may be used as reference knowledge base for biological interpretation of large-scale datasets generated by sequencing and other high-throughput experimental technologies.
View all literature mentionsOpen source software package for statistical programming language R to create plots based on grammar of graphics. Used for data visualization to break up graphs into semantic components such as scales and layers.
View all literature mentionsTHIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software tool for the rapid annotation of prokaryotic genomes. It produces GFF3, GBK and SQN files that are ready for editing in Sequin and ultimately submitted to Genbank/DDJB/ENA. A typical 4 Mbp genome can be fully annotated in less than 10 minutes on a quad-core computer, and scales well to 32 core SMP systems.
View all literature mentionsA pathway-based graphical interface for navigating the glycoenzyme database. The goal of the project is to define the paradigms by which carbohydrate binding proteins function in cellular communication. These pages are divided into six categories: -Glycosphingolipid: Sub-categories are Isogloboseries, Globoseries, Neo-lactoseries, Lactoseries and Ganglioseries - N-linked: Sub-categories are High-mannose, Hybrid and Complex -Mucin -Terminal Core 1 -Other O-linked -Terminal All: Includes all potential terminal structures for each glycan category
View all literature mentions