Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
How the human brain differs from those of non-human primates is largely unknown and the complex drivers underlying such differences at the genomic level remain unclear. In this study, we selected 243 brain-related genes, based on Gene Ontology, and identified 184,113 DNaseI hypersensitive sites (DHSs) within their regulatory regions. To performed comprehensive evolutionary analyses, we set strict filtering criteria for alignment quality and filtered 39,132 DHSs for inclusion in the investigation and found that 2,397 (~6%) exhibited evidence of accelerated evolution (aceDHSs), which was a much higher proportion that DHSs genome-wide. Target genes predicted to be regulated by brain-aceDHSs were functionally enriched for brain development and exhibited differential expression between human and chimpanzee. Alignments indicated 61 potential human-specific transcription factor binding sites in brain-aceDHSs, including for CTCF, FOXH1, and FOXQ1. Furthermore, based on GWAS, Hi-C, and eQTL data, 16 GWAS SNPs, and 82 eQTL SNPs were in brain-aceDHSs that regulate genes related to brain development or disease. Among these brain-aceDHSs, we confirmed that one enhanced the expression of GPR133, using CRISPR-Cas9 and western blotting. The GPR133 gene is associated with glioblastoma, indicating that SNPs within DHSs could be related to brain disorders. These findings suggest that brain-related gene regulatory regions are under adaptive evolution and contribute to the differential expression profiles among primates, providing new insights into the genetic basis of brain phenotypes or disorders between humans and other primates.
Pubmed ID: 30930929
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Bioinformatics resource system including web server and web service for functional annotation and enrichment analyses of gene lists. Consists of comprehensive knowledgebase and set of functional analysis tools. Includes gene centered database integrating heterogeneous gene annotation resources to facilitate high throughput gene functional analysis.
View all literature mentionsCollection of genome databases for vertebrates and other eukaryotic species with DNA and protein sequence search capabilities. Used to automatically annotate genome, integrate this annotation with other available biological data and make data publicly available via web. Ensembl tools include BLAST, BLAT, BioMart and the Variant Effect Predictor (VEP) for all supported species.
View all literature mentionsDatabase for genomes that have been completely sequenced, have active research community to contribute gene-specific information, or that are scheduled for intense sequence analysis. Includes nomenclature, map location, gene products and their attributes, markers, phenotypes, and links to citations, sequences, variation details, maps, expression, homologs, protein domains and external databases. All entries follow NCBI's format for data collections. Content of Entrez Gene represents result of curation and automated integration of data from NCBI's Reference Sequence project (RefSeq), from collaborating model organism databases, and from many other databases available from NCBI. Records are assigned unique, stable and tracked integers as identifiers. Content is updated as new information becomes available.
View all literature mentionsOpen source database of curated, non-redundant set of profiles derived from published collections of experimentally defined transcription factor binding sites for multicellular eukaryotes. Consists of open data access, non-redundancy and quality. JASPAR CORE is smaller set that is non-redundant and curated. Collection of transcription factor DNA-binding preferences, modeled as matrices. These can be converted into Position Weight Matrices (PWMs or PSSMs), used for scanning genomic sequences. Web interface for browsing, searching and subset selection, online sequence analysis utility and suite of programming tools for genome-wide and comparative genomic analysis of regulatory regions. New functions include clustering of matrix models by similarity, generation of random matrices by sampling from selected sets of existing models and a language-independent Web Service applications programming interface for matrix retrieval.
View all literature mentionsWeb based gene set analysis toolkit designed for functional genomic, proteomic, and large-scale genetic studies from which large number of gene lists (e.g. differentially expressed gene sets, co-expressed gene sets etc) are continuously generated. WebGestalt incorporates information from different public resources and provides a way for biologists to make sense out of gene lists. This version of WebGestalt supports eight organisms, including human, mouse, rat, worm, fly, yeast, dog, and zebrafish.
View all literature mentionsA suite of tools to address common questions raised in genomic studies - mostly with regard to overlap and proximity relationships between data sets.
View all literature mentions