Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
The adaptive radiations of East African cichlid fish in the Great Lakes Victoria, Malawi, and Tanganyika are well known for their diversity and repeatedly evolved phenotypes. Convergent evolution of melanic horizontal stripes has been linked to a single locus harboring the gene agouti-related peptide 2 (agrp2). However, where and when the causal variants underlying this trait evolved and how they drove phenotypic divergence remained unknown. To test the alternative hypotheses of standing genetic variation versus de novo mutations (independently originating in each radiation), we searched for shared signals of genomic divergence at the agrp2 locus. Although we discovered similar signatures of differentiation at the locus level, the haplotypes associated with stripe patterns are surprisingly different. In Lake Malawi, the highest associated alleles are located within and close to the 5' untranslated region of agrp2 and likely evolved through recent de novo mutations. In the younger Lake Victoria radiation, stripes are associated with two intronic regions overlapping with a previously reported cis-regulatory interval. The origin of these segregating haplotypes predates the Lake Victoria radiation because they are also found in more basal riverine and Lake Kivu species. This suggests that both segregating haplotypes were present as standing genetic variation at the onset of the Lake Victoria adaptive radiation with its more than 500 species and drove phenotypic divergence within the species flock. Therefore, both new (Lake Malawi) and ancient (Lake Victoria) allelic variation at the same locus fueled rapid and convergent phenotypic evolution.
Pubmed ID: 32941629
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
A C++ library for parsing and manipulating Variant Call Format (VCF) files, and many command-line utilities. The API provides a quick and extremely permissive method to read and write VCF files. Extensions and applications of the library provided in the included utilities (*.cpp) comprise the vast bulk of the library's utility for most users.
View all literature mentionsSoftware package for working with VCF files. Used to provide easily accessible methods for working with complex genetic variation data in the form of VCF files.Implements various utilities for processing Variant Call Format files, including validation, merging, comparing. Provides general Perl API.
View all literature mentionsA software package to analyze next-generation resequencing data. The toolkit offers a wide variety of tools, with a primary focus on variant discovery and genotyping as well as strong emphasis on data quality assurance. Its robust architecture, powerful processing engine and high-performance computing features make it capable of taking on projects of any size. This software library makes writing efficient analysis tools using next-generation sequencing data very easy, and second it's a suite of tools for working with human medical resequencing projects such as 1000 Genomes and The Cancer Genome Atlas. These tools include things like a depth of coverage analyzers, a quality score recalibrator, a SNP/indel caller and a local realigner. (entry from Genetic Analysis Software)
View all literature mentionsOriginal SAMTOOLS package has been split into three separate repositories including Samtools, BCFtools and HTSlib. Samtools for manipulating next generation sequencing data used for reading, writing, editing, indexing,viewing nucleotide alignments in SAM,BAM,CRAM format. BCFtools used for reading, writing BCF2,VCF, gVCF files and calling, filtering, summarising SNP and short indel sequence variants. HTSlib used for reading, writing high throughput sequencing data.
View all literature mentionsModel organism database that serves as central repository and web-based resource for zebrafish genetic, genomic, phenotypic and developmental data. Data represented are derived from three primary sources: curation of zebrafish publications, individual research laboratories and collaborations with bioinformatics organizations. Data formats include text, images and graphical representations.Serves as primary community database resource for laboratory use of zebrafish. Developed and supports integrated zebrafish genetic, genomic, developmental and physiological information and link this information extensively to corresponding data in other model organism and human databases.
View all literature mentionsOpen source database of curated, non-redundant set of profiles derived from published collections of experimentally defined transcription factor binding sites for multicellular eukaryotes. Consists of open data access, non-redundancy and quality. JASPAR CORE is smaller set that is non-redundant and curated. Collection of transcription factor DNA-binding preferences, modeled as matrices. These can be converted into Position Weight Matrices (PWMs or PSSMs), used for scanning genomic sequences. Web interface for browsing, searching and subset selection, online sequence analysis utility and suite of programming tools for genome-wide and comparative genomic analysis of regulatory regions. New functions include clustering of matrix models by similarity, generation of random matrices by sampling from selected sets of existing models and a language-independent Web Service applications programming interface for matrix retrieval.
View all literature mentionsA graphical viewer of phylogenetic trees and a program for producing publication-ready figures. It is designed to display summarized and annotated trees produced by BEAST.
View all literature mentionsWeb phylogeny server based on the maximum-likelihood principle.
View all literature mentionsData aggregate that compiles results from bioinformatics analyses across multiple samples into a single report. It is written in Python.
View all literature mentionsSoftware tool used to carry out statistical selection of best-fit models of nucleotide substitution without the aid of PAUP*. It implements five different model selection strategies: hierarchical and dynamical likelihood ratio tests, Akaike and Bayesian information criteria, and a decision theory method. It also provides estimates of model selection uncertainty, parameter importances, and model-averaged parameter estimates.
View all literature mentionsJava toolset for working with next generation sequencing data in the BAM format.
View all literature mentions