Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
This study describes the structure of DNA polymerase I from Thermus phage G20c, termed PolI_G20c. This is the first structure of a DNA polymerase originating from a group of related thermophilic bacteriophages infecting Thermus thermophilus, including phages G20c, TSP4, P74-26, P23-45 and phiFA and the novel phage Tth15-6. Sequence and structural analysis of PolI_G20c revealed a 3'-5' exonuclease domain and a DNA polymerase domain, and activity screening confirmed that both domains were functional. No functional 5'-3' exonuclease domain was present. Structural analysis also revealed a novel specific structure motif, here termed SβαR, that was not previously identified in any polymerase belonging to the DNA polymerases I (or the DNA polymerase A family). The SβαR motif did not show any homology to the sequences or structures of known DNA polymerases. The exception was the sequence conservation of the residues in this motif in putative DNA polymerases encoded in the genomes of a group of thermophilic phages related to Thermus phage G20c. The structure of PolI_G20c was determined with the aid of another structure that was determined in parallel and was used as a model for molecular replacement. This other structure was of a 3'-5' exonuclease termed ExnV1. The cloned and expressed gene encoding ExnV1 was isolated from a thermophilic virus metagenome that was collected from several hot springs in Iceland. The structure of ExnV1, which contains the novel SβαR motif, was first determined to 2.19 Å resolution. With these data at hand, the structure of PolI_G20c was determined to 2.97 Å resolution. The structures of PolI_G20c and ExnV1 are most similar to those of the Klenow fragment of DNA polymerase I (PDB entry 2kzz) from Escherichia coli, DNA polymerase I from Geobacillus stearothermophilus (PDB entry 1knc) and Taq polymerase (PDB entry 1bgx) from Thermus aquaticus.
Pubmed ID: 36322421
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
NIH genetic sequence database that provides annotated collection of all publicly available DNA sequences for almost 280 000 formally described species (Jan 2014) .These sequences are obtained primarily through submissions from individual laboratories and batch submissions from large-scale sequencing projects, including whole-genome shotgun (WGS) and environmental sampling projects. Most submissions are made using web-based BankIt or standalone Sequin programs, and GenBank staff assigns accession numbers upon data receipt. It is part of International Nucleotide Sequence Database Collaboration and daily data exchange with European Nucleotide Archive (ENA) and DNA Data Bank of Japan (DDBJ) ensures worldwide coverage. GenBank is accessible through NCBI Entrez retrieval system, which integrates data from major DNA and protein sequence databases along with taxonomy, genome, mapping, protein structure and domain information, and biomedical journal literature via PubMed. BLAST provides sequence similarity searches of GenBank and other sequence databases. Complete bimonthly releases and daily updates of GenBank database are available by FTP.
View all literature mentionsService providing functional analysis of proteins by classifying them into families and predicting domains and important sites. They combine protein signatures from a number of member databases into a single searchable resource, capitalizing on their individual strengths to produce a powerful integrated database and diagnostic tool. This integrated database of predictive protein signatures is used for the classification and automatic annotation of proteins and genomes. InterPro classifies sequences at superfamily, family and subfamily levels, predicting the occurrence of functional domains, repeats and important sites. InterPro adds in-depth annotation, including GO terms, to the protein signatures. You can access the data programmatically, via Web Services. The member databases use a number of approaches: # ProDom: provider of sequence-clusters built from UniProtKB using PSI-BLAST. # PROSITE patterns: provider of simple regular expressions. # PROSITE and HAMAP profiles: provide sequence matrices. # PRINTS provider of fingerprints, which are groups of aligned, un-weighted Position Specific Sequence Matrices (PSSMs). # PANTHER, PIRSF, Pfam, SMART, TIGRFAMs, Gene3D and SUPERFAMILY: are providers of hidden Markov models (HMMs). Your contributions are welcome. You are encouraged to use the ''''Add your annotation'''' button on InterPro entry pages to suggest updated or improved annotation for individual InterPro entries.
View all literature mentionsHarvester is a Web-based tool that bulk-collects bioinformatic data on human proteins from various databases and prediction servers. It is a meta search engine for gene and protein information. It searches 16 major databases and prediction servers and combines the results on pregenerated HTML pages. In this way Harvester can provide comprehensive gene-protein information from different servers in a convenient and fast manner. As full text meta search engine, similar to Google trade mark, Harvester allows screening of the whole genome proteome for current protein functions and predictions in a few seconds. With Harvester it is now possible to compare and check the quality of different database entries and prediction algorithms on a single page. Sponsors: This work has been supported by the BMBF with grants 01GR0101 and 01KW0013.
View all literature mentionsCommercial vendor and service provider of laboratory reagents and antibodies. Supplier of scientific instrumentation, reagents and consumables, and software services.
View all literature mentionsPortal which provides access to scientific databases and software tools (i.e., resources) in different areas of life sciences including proteomics, genomics, phylogeny, systems biology, population genetics, transcriptomics etc. It contains resources from many different SIB groups as well as external institutions.
View all literature mentions