Searching the Resource Information Network

Our searching services are busy right now. Please try again later

  • Register
X
Forgot Password

If you have forgotten your password you can enter your email here and get a temporary password sent to your email.

X

Leaving Community

Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.

No
Yes

Complete Genome Sequencing and Comparative Genome Characterization of Lactobacillus johnsonii ZLJ010, a Potential Probiotic With Health-Promoting Properties.

Wei Zhang | Jing Wang | Dongyan Zhang | Hui Liu | Sixin Wang | Yamin Wang | Haifeng Ji
Frontiers in genetics | 2019

Lactobacillus johnsonii ZLJ010 is a probiotic strain isolated from the feces of a healthy sow and has putative health-promoting properties. To determine the molecular basis underlying the probiotic potential of ZLJ010 and the genes involved in the same, complete genome sequencing and comparative genome analysis with L. johnsonii ZLJ010 were performed. The ZLJ010 genome was found to contain a single circular chromosome of 1,999,879 bp with a guanine-cytosine (GC) content of 34.91% and encoded 18 ribosomal RNA (rRNA) genes and 77 transfer RNA (tRNA) genes. From among the 1,959 protein coding sequences (CDSs), genes known to confer probiotic properties were identified, including genes related to stress adaptation, biosynthesis, metabolism, transport of amino acid, secretion, and the defense machinery. ZLJ010 lacked complete or partial biosynthetic pathways for amino acids but was predicted to compensate for this with an enhanced transport system and some unique amino acid permeases and peptidases that allow it to acquire amino acids and other precursors exogenously. The comparative genomic analysis of L. johnsonii ZLP001 and seven other available L. johnsonii strains, including L. johnsonii NCC533, FI9785, DPC6026, N6.2, BS15, UMNLJ22, and PF01, revealed 2,732 pan-genome orthologous gene clusters and 1,324 core-genome orthologous gene clusters. Phylogenomic analysis based on 1,288 single copy genes showed that ZLJ010 had a closer relationship with the BS15 from yogurt and DPC6026 from the porcine intestinal tract but was located on a relatively standalone branch. The number of clusters of unique, strain-specific genes ranged from 42 to 185. A total of 219 unique genes present in the genome of L. johnsonii ZLJ010 primarily encoded proteins that are putatively involved in replication, recombination and repair, defense mechanisms, transcription, amino acid transport and metabolism, and carbohydrate transport and metabolism. Two unique prophages were predicted in the ZLJ010 genome. The present study helps us understand the ability of L. johnsonii ZLJ010 to better adapt to the gut environment and also its probiotic functionalities.

Pubmed ID: 31552103

Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.

This is a list of tools and resources that we have found mentioned in this publication.


GenBank (tool)

RRID:SCR_002760

NIH genetic sequence database that provides annotated collection of all publicly available DNA sequences for almost 280 000 formally described species (Jan 2014) .These sequences are obtained primarily through submissions from individual laboratories and batch submissions from large-scale sequencing projects, including whole-genome shotgun (WGS) and environmental sampling projects. Most submissions are made using web-based BankIt or standalone Sequin programs, and GenBank staff assigns accession numbers upon data receipt. It is part of International Nucleotide Sequence Database Collaboration and daily data exchange with European Nucleotide Archive (ENA) and DNA Data Bank of Japan (DDBJ) ensures worldwide coverage. GenBank is accessible through NCBI Entrez retrieval system, which integrates data from major DNA and protein sequence databases along with taxonomy, genome, mapping, protein structure and domain information, and biomedical journal literature via PubMed. BLAST provides sequence similarity searches of GenBank and other sequence databases. Complete bimonthly releases and daily updates of GenBank database are available by FTP.

View all literature mentions

RAxML (tool)

RRID:SCR_006086

Software program for phylogenetic analyses of large datasets under maximum likelihood.

View all literature mentions

Promega (tool)

RRID:SCR_006724

An Antibody supplier

View all literature mentions

Rfam (tool)

RRID:SCR_007891

The Rfam database is a collection of RNA families, each represented by multiple sequence alignments, consensus secondary structures and covariance models (CMs). The families in Rfam break down into three broad functional classes: Non-coding RNA genes, structured cis-regulatory elements and self-splicing RNAs. Typically these functional RNAs often have a conserved secondary structure which may be better preserved than the RNA sequence. The CMs used to describe each family are a slightly more complicated relative of the profile hidden Markov models (HMMs) used by Pfam. CMs can simultaneously model RNA sequence and the structure in an elegant and accurate fashion. Rfam is also available via FTP. You can find data in Rfam in various ways... * Analyze your RNA sequence for Rfam matches * View Rfam family annotation and alignments * View Rfam clan details * Query Rfam by keywords * Fetch families or sequences by NCBI taxonomy * Enter any type of accession or ID to jump to the page for a Rfam family, sequence or genome

View all literature mentions

Thermo Fisher Scientific (tool)

RRID:SCR_008452

Commercial vendor and service provider of laboratory reagents and antibodies. Supplier of scientific instrumentation, reagents and consumables, and software services.

View all literature mentions

SOAPdenovo (tool)

RRID:SCR_010752

THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 24,2023. Software tool for de novo assembly of human genomes with massively parallel short read sequencing.Short-read assembly method that can build de novo draft assembly for human sized genomes.Software package for assembling short oligonucleotide into contigs and scaffolds.

View all literature mentions

Circos (tool)

RRID:SCR_011798

A software package for visualizing data and information. It visualizes data in a circular layout - this makes Circos ideal for exploring relationships between objects or positions.

View all literature mentions

MAFFT (tool)

RRID:SCR_011811

Software package as multiple alignment program for amino acid or nucleotide sequences. Can align up to 500 sequences or maximum file size of 1 MB. First version of MAFFT used algorithm based on progressive alignment, in which sequences were clustered with help of Fast Fourier Transform. Subsequent versions have added other algorithms and modes of operation, including options for faster alignment of large numbers of sequences, higher accuracy alignments, alignment of non-coding RNA sequences, and addition of new sequences to existing alignments.

View all literature mentions

Glimmer (tool)

RRID:SCR_011931

A software system for finding genes in microbial DNA, especially the genomes of bacteria, archaea, and viruses.

View all literature mentions

KEGG (tool)

RRID:SCR_012773

Integrated database resource consisting of 16 main databases, broadly categorized into systems information, genomic information, and chemical information. In particular, gene catalogs in completely sequenced genomes are linked to higher-level systemic functions of cell, organism, and ecosystem. Analysis tools are also available. KEGG may be used as reference knowledge base for biological interpretation of large-scale datasets generated by sequencing and other high-throughput experimental technologies.

View all literature mentions

CAZy- Carbohydrate Active Enzyme (tool)

RRID:SCR_012909

Database that describes the families of structurally-related catalytic and carbohydrate-binding modules (or functional domains) of enzymes that degrade, modify, or create glycosidic bonds. This specialist database is dedicated to the display and analysis of genomic, structural and biochemical information on Carbohydrate-Active Enzymes (CAZymes). CAZy data are accessible either by browsing sequence-based families or by browsing the content of genomes in carbohydrate-active enzymes. New genomes are added regularly shortly after they appear in the daily releases of GenBank. New families are created based on published evidence for the activity of at least one member of the family and all families are regularly updated, both in content and in description. An original aspect of the CAZy database is its attempt to cover all carbohydrate-active enzymes across organisms and across subfields of glycosciences. One can search for CAZY Family pages using the Protein Accession (Genpept Accession, Uniprot Accession or PDB ID), Cazy family name or EC number. In addition, genomes can be searched using the NCBI TaxID. This search can be complemented by Google-based searches on the CAZy site.

View all literature mentions

SignalP (tool)

RRID:SCR_015644

Web application for prediction of the presence and location of signal peptide cleavage sites in amino acid sequences from different organisms. The method incorporates a prediction of cleavage sites and a signal peptide/non-signal peptide prediction based on a combination of several artificial neural networks.

View all literature mentions