Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
Whole genome analysis of a novel species of enterococci, Enterococcus lacertideformus, causing multi-systemic and invariably fatal disease in critically endangered Christmas Island reptiles was undertaken to determine the genetic elements and potential mechanisms conferring its pathogenic nature, biofilm-forming capabilities, immune recognition avoidance, and inability to grow in vitro. Comparative genomic analyses with related and clinically significant enterococci were further undertaken to infer the evolutionary history of the bacterium and identify genes both novel and absent. The genome had a G + C content of 35.1%, consisted of a circular chromosome, no plasmids, and was 2,419,934 bp in length (2,321 genes, 47 tRNAs, and 13 rRNAs). Multi-locus sequence typing (MLST), and single nucleotide polymorphism (SNP) analysis of multiple E. lacertideformus samples revealed they were effectively indistinguishable from one another and highly clonal. E. lacertideformus was found to be located within the Enterococcus faecium species clade and was closely related to Enterococcus villorum F1129D based on 16S rDNA and MLST house-keeping gene analysis. Antimicrobial resistance (DfreE, EfrB, tetM, bcrRABD, and sat4) and virulence genes (Fss3 and ClpP), and genes conferring tolerance to metals and biocides (n = 9) were identified. The detection of relatively few genes encoding antimicrobial resistance and virulence indicates that this bacterium may have had no exposure to recently developed and clinically significant antibiotics. Genes potentially imparting beneficial functional properties were identified, including prophages, insertion elements, integrative conjugative elements, and genomic islands. Functional CRISPR-Cas arrays, and a defective prophage region were identified in the genome. The study also revealed many genomic loci unique to E. lacertideformus which contained genes enriched in cell wall/membrane/envelop biogenesis, and carbohydrate metabolism and transport functionality. This finding and the detection of putative enterococcal biofilm determinants (EfaAfs, srtC, and scm) may underpin the novel biofilm phenotype observed for this bacterium. Comparative analysis of E. lacertideformus with phylogenetically related and clinically significant enterococci (E. villorum F1129D, Enterococcus hirae R17, E. faecium AUS0085, and Enterococcus faecalis OG1RF) revealed an absence of genes (n = 54) in E. lacertideformus, that encode metabolic functionality, which potentially hinders nutrient acquisition and/or utilization by the bacterium and precludes growth in vitro. These data provide genetic insights into the previously determined phenotype and pathogenic nature of the bacterium.
Pubmed ID: 33737921
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Quality assessment software tool for evaluating and comparing genome assemblies. It works both with and without a given reference genome. It produces many reports, summary tables and plots.
View all literature mentionsWeb application to search nucleotide databases using a nucleotide query. Algorithms: blastn, megablast, discontiguous megablast.
View all literature mentionsWeb application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.
View all literature mentionsNIH genetic sequence database that provides annotated collection of all publicly available DNA sequences for almost 280 000 formally described species (Jan 2014) .These sequences are obtained primarily through submissions from individual laboratories and batch submissions from large-scale sequencing projects, including whole-genome shotgun (WGS) and environmental sampling projects. Most submissions are made using web-based BankIt or standalone Sequin programs, and GenBank staff assigns accession numbers upon data receipt. It is part of International Nucleotide Sequence Database Collaboration and daily data exchange with European Nucleotide Archive (ENA) and DNA Data Bank of Japan (DDBJ) ensures worldwide coverage. GenBank is accessible through NCBI Entrez retrieval system, which integrates data from major DNA and protein sequence databases along with taxonomy, genome, mapping, protein structure and domain information, and biomedical journal literature via PubMed. BLAST provides sequence similarity searches of GenBank and other sequence databases. Complete bimonthly releases and daily updates of GenBank database are available by FTP.
View all literature mentionsA graphical viewer of phylogenetic trees and a program for producing publication-ready figures. It is designed to display summarized and annotated trees produced by BEAST.
View all literature mentionsSoftware to: view dicom files and assemble them into 3D volumes. View and convert between Analyze, Nifti, and Interfile. Classify and organize dicoms and 3D volumes using metadata. Search and report on a collection of scans.
View all literature mentionsSoftware package for sequence alignment, assembly and analysis. Integrated and extendable desktop software platform for organization and analysis of sequence data. Bioinformatics software platform packed with molecular biology and sequence analysis tools.
View all literature mentionsSoftware package as multiple alignment program for amino acid or nucleotide sequences. Can align up to 500 sequences or maximum file size of 1 MB. First version of MAFFT used algorithm based on progressive alignment, in which sequences were clustered with help of Fast Fourier Transform. Subsequent versions have added other algorithms and modes of operation, including options for faster alignment of large numbers of sequences, higher accuracy alignments, alignment of non-coding RNA sequences, and addition of new sequences to existing alignments.
View all literature mentionsSoftware Java pipeline for trimming tasks for Illumina paired end and single ended data. Flexible Trimmer for Illumina Sequence Data. Pair aware preprocessing tool optimized for Illumina next generation sequencing data. Includes several processing steps for read trimming and filtering. Operating systems Unix/Linux, Mac OS, Windows.
View all literature mentionsQuality control software that perform checks on raw sequence data coming from high throughput sequencing pipelines. This software also provides a modular set of analyses which can give a quick impression of the quality of the data prior to further analysis.
View all literature mentionsWeb phylogeny server based on the maximum-likelihood principle.
View all literature mentionsSoftware tool as a short read aligner for DNA and RNA seq data. Used for large genomes with millions of scaffolds. Can align reads from Illumina, PacBio, 454, Sanger, Ion Torrent, Nanopore. Fast and accurate, particularly with highly mutated genomes or reads with long indels, even whole gene deletions over 100kbp long. It has no upper limit to genome size or number of contigs. Written in Java, can run on any platform.
View all literature mentionsSoftware tool as Next Generation Sequencing assembler. Optimized for metagenomes, but also works well on generic single genome assembly (small or mammalian size) and single cell assembly. Can assemble genome sequences from metagenomic datasets of hundreds of Giga base-pairs in time and memory efficient manner on single server.
View all literature mentionsSoftware tool for sequence mapping.The next version of BWA-MEM. Used for aligning sequencing reads against large reference genome.
View all literature mentions