Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
Marine Fungi are potent secondary metabolite producers. However, limited genetic information are available their biosynthetic gene clusters (BGCs) and their biotechnological applications. To overcome this lack of information, herein, we used next-generation sequencing methods for genome sequencing of two marine fungi, isolated from the German Wadden Sea, namely Calcarisporium sp. KF525 and Pestalotiopsis sp. KF079. The assembled genome size of the marine isolate Calcarisporium sp. KF525 is about 36.8 Mb with 60 BGCs, while Pestalotiopsis sp. KF079 has a genome size of 47.5 Mb harboring 67 BGCs. Of all BGCs, 98% and 97% are novel clusters of Calcarisporium sp. and Pestalotiopsis sp., respectively. Only few of the BGCs were found to be expressed under laboratory conditions by RNA-seq analysis. The vast majority of all BGCs were found to be novel and unique for these two marine fungi. Along with a description of the identified gene clusters, we furthermore present important genomic features and life-style properties of these two fungi. The two novel fungal genomes provide a plethora of new BGCs, which may have biotechnological applications in the future, for example as novel drugs. The genomic characterizations will provide assistance in future genetics and genomic analyses of marine fungi.
Pubmed ID: 29976990
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Database of biological data related to a single initiative, originating from a single organization or from a consortium. A BioProject record provides users a single place to find links to the diverse data types generated for that project. It is a searchable collection of complete and incomplete (in-progress) large-scale sequencing, assembly, annotation, and mapping projects for cellular organisms. Submissions are supported by a web-based Submission Portal. The database facilitates organization and classification of project data submitted to NCBI, EBI and DDBJ databases that captures descriptive information about research projects that result in high volume submissions to archival databases, ties together related data across multiple archives and serves as a central portal by which to inform users of data availability. BioProject records link to corresponding data stored in archival repositories. The BioProject resource is a redesigned, expanded, replacement of the NCBI Genome Project resource. The redesign adds tracking of several data elements including more precise information about a project''''s scope, material, and objectives. Genome Project identifiers are retained in the BioProject as the ID value for a record, and an Accession number has been added. Database content is exchanged with other members of the International Nucleotide Sequence Database Collaboration (INSDC). BioProject is accessible via FTP.
View all literature mentionsDatabase containing descriptions of biological source materials used in experimental assays. Sources include: GenBank, Sequence Read Archive (SRA), Coriell, ATCC. Submissions are supported by a web-based Submission Portal that guides users through a series of forms for input of rich metadata describing their samples. As the capacity and complexity of biological data sets expands, databases face new challenges in ensuring that the information is adequately organized and described. The NCBI BioSample database is being developed to help address the challenges by providing the means by which data generators can organize and describe a broad range of sample types, and link to corresponding sets of experimental data in archival databases.
View all literature mentionsService providing functional analysis of proteins by classifying them into families and predicting domains and important sites. They combine protein signatures from a number of member databases into a single searchable resource, capitalizing on their individual strengths to produce a powerful integrated database and diagnostic tool. This integrated database of predictive protein signatures is used for the classification and automatic annotation of proteins and genomes. InterPro classifies sequences at superfamily, family and subfamily levels, predicting the occurrence of functional domains, repeats and important sites. InterPro adds in-depth annotation, including GO terms, to the protein signatures. You can access the data programmatically, via Web Services. The member databases use a number of approaches: # ProDom: provider of sequence-clusters built from UniProtKB using PSI-BLAST. # PROSITE patterns: provider of simple regular expressions. # PROSITE and HAMAP profiles: provide sequence matrices. # PRINTS provider of fingerprints, which are groups of aligned, un-weighted Position Specific Sequence Matrices (PSSMs). # PANTHER, PIRSF, Pfam, SMART, TIGRFAMs, Gene3D and SUPERFAMILY: are providers of hidden Markov models (HMMs). Your contributions are welcome. You are encouraged to use the ''''Add your annotation'''' button on InterPro entry pages to suggest updated or improved annotation for individual InterPro entries.
View all literature mentionsDatabase that describes the families of structurally-related catalytic and carbohydrate-binding modules (or functional domains) of enzymes that degrade, modify, or create glycosidic bonds. This specialist database is dedicated to the display and analysis of genomic, structural and biochemical information on Carbohydrate-Active Enzymes (CAZymes). CAZy data are accessible either by browsing sequence-based families or by browsing the content of genomes in carbohydrate-active enzymes. New genomes are added regularly shortly after they appear in the daily releases of GenBank. New families are created based on published evidence for the activity of at least one member of the family and all families are regularly updated, both in content and in description. An original aspect of the CAZy database is its attempt to cover all carbohydrate-active enzymes across organisms and across subfields of glycosciences. One can search for CAZY Family pages using the Protein Accession (Genpept Accession, Uniprot Accession or PDB ID), Cazy family name or EC number. In addition, genomes can be searched using the NCBI TaxID. This search can be complemented by Google-based searches on the CAZy site.
View all literature mentionsA company that provides a variety of next generation sequencing services. The company provides researchers with whole genome resequencing, exome sequencing, targeted sequencing, transcriptomics, and epigenome sequencing.
View all literature mentions