Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
Acute lymphoblastic leukaemia (ALL) is the most common paediatric malignancy. Genome-wide association studies have shown variation at 14q11.2 influences ALL risk. We sought to decipher causal variant(s) at 14q11.2 and the mechanism of tumorigenesis. We show rs2239630 G>A resides in the promoter of the CCAT enhancer-binding protein epsilon (CEBPE) gene. The rs2239630-A risk allele is associated with increased promotor activity and CEBPE expression. Depletion of CEBPE in ALL cells reduces cell growth, correspondingly CEBPE binds to the promoters of electron transport and energy generation genes. RNA-seq in CEBPE depleted cells demonstrates CEBPE regulates the expression of genes involved in B-cell development (IL7R), apoptosis (BCL2), and methotrexate resistance (RASS4L). CEBPE regulated genes significantly overlapped in CEBPE depleted cells, ALL blasts and IGH-CEBPE translocated ALL. This suggests CEBPE regulates a similar set of genes in each, consistent with a common biological mechanism of leukemogenesis for rs2239630 associated and CEBPE translocated ALL. Finally, we map IGH-CEBPE translocation breakpoints in two cases, implicating RAG recombinase activity in their formation.
Pubmed ID: 29977016
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Non-profit plasmid repository dedicated to helping scientists around the world share high-quality plasmids. Facilitates archiving and distributing DNA-based research reagents and associated data to scientists worldwide. Repository contains over 65,000 plasmids, including special collections on CRISPR, fluorescent proteins, and ready-to-use viral preparations. There is no cost for scientists to deposit plasmids, which saves time and money associated with shipping plasmids themselves. All plasmids are fully sequenced for validation and sequencing data is openly available. We handle the appropriate Material Transfer Agreements (MTA) with institutions, facilitating open exchange and offering intellectual property and liability protection for depositing scientists. Furthermore, we curate free educational resources for the scientific community including a blog, eBooks, video protocols, and detailed molecular biology resources.
View all literature mentionsA software package for the mapping of short reads from illumina sequencing machines onto a reference genome. It''s recommended for most workflows, including those for genomic resequencing, RNA-Seq and Chip-seq. Stampy excels in the mapping of reads containing that contain sequence variation relative to the reference, in particular for those containing insertions or deletions. It can map reads from a highly divergent species to a reference genome for instance. Stampy achieves high sensitivity and speed by using a fast hashing algorithm and a detailed statistical model. Stampy has the following features: * Maps single, paired-end and mate pair Illumina reads to a reference genome * Fast: about 20 Gbase per hour in hybrid mode (using BWA) * Low memory footprint: 2.7 Gb shared memory for a 3Gbase genome * High sensitivity for indels and divergent reads, up to 10-15% * Low mapping bias for reads with SNPs * Well calibrated mapping quality scores * Input: Fastq and Fasta; gzipped or plain * Output: SAM, Maq''s map file * Optionally calculates per-base alignment posteriors * Optionally processes part of the input * Handles reads of up to 4500 bases
View all literature mentionsA next-generation web-based application that aims to provide an integrated solution for both visualization and analysis of deep-sequencing data, along with simple access to public datasets.
View all literature mentionsEncyclopedia of DNA elements consisting of list of functional elements in human genome, including elements that act at protein and RNA levels, and regulatory elements that control cells and circumstances in which gene is active. Enables scientific and medical communities to interpret role of human genome in biology and disease. Provides identification of common cell types to facilitate integrative analysis and new experimental technologies based on high-throughput sequencing. Genome Browser containing ENCODE and Epigenomics Roadmap data. Data are available for entire human genome.
View all literature mentionsSoftware for single-cell flow cytometry analysis. Its functions include management, display, manipulation, analysis and publication of the data stream produced by flow and mass cytometers.
View all literature mentionsA dataset containing the full genomic sequence of 1,700 individuals, freely available for research use. The 1000 Genomes Project is an international research effort coordinated by a consortium of 75 companies and organizations to establish the most detailed catalogue of human genetic variation. The project has grown to 200 terabytes of genomic data including DNA sequenced from more than 1,700 individuals that researchers can now access on AWS for use in disease research free of charge. The dataset containing the full genomic sequence of 1,700 individuals is now available to all via Amazon S3. The data can be found at: http://s3.amazonaws.com/1000genomes The 1000 Genomes Project aims to include the genomes of more than 2,662 individuals from 26 populations around the world, and the NIH will continue to add the remaining genome samples to the data collection this year. Public Data Sets on AWS provide a centralized repository of public data hosted on Amazon Simple Storage Service (Amazon S3). The data can be seamlessly accessed from AWS services such Amazon Elastic Compute Cloud (Amazon EC2) and Amazon Elastic MapReduce (Amazon EMR), which provide organizations with the highly scalable compute resources needed to take advantage of these large data collections. AWS is storing the public data sets at no charge to the community. Researchers pay only for the additional AWS resources they need for further processing or analysis of the data. All 200 TB of the latest 1000 Genomes Project data is available in a publicly available Amazon S3 bucket. You can access the data via simple HTTP requests, or take advantage of the AWS SDKs in languages such as Ruby, Java, Python, .NET and PHP. Researchers can use the Amazon EC2 utility computing service to dive into this data without the usual capital investment required to work with data at this scale. AWS also provides a number of orchestration and automation services to help teams make their research available to others to remix and reuse. Making the data available via a bucket in Amazon S3 also means that customers can crunch the information using Hadoop via Amazon Elastic MapReduce, and take advantage of the growing collection of tools for running bioinformatics job flows, such as CloudBurst and Crossbow.
View all literature mentionsSoftware tools for Motif Discovery and next-gen sequencing analysis. Used for analyzing ChIP-Seq, GRO-Seq, RNA-Seq, DNase-Seq, Hi-C and numerous other types of functional genomics sequencing data sets. Collection of command line programs for unix style operating systems written in Perl and C++.
View all literature mentionsSoftware Java pipeline for trimming tasks for Illumina paired end and single ended data. Flexible Trimmer for Illumina Sequence Data. Pair aware preprocessing tool optimized for Illumina next generation sequencing data. Includes several processing steps for read trimming and filtering. Operating systems Unix/Linux, Mac OS, Windows.
View all literature mentionsBioconductor software package for Empirical analysis of Digital Gene Expression data in R. Used for differential expression analysis of RNA-seq and digital gene expression data with biological replication.
View all literature mentionsA computer program for phasing observed genotypes and imputing missing genotypes.
View all literature mentionsConsortium to build comprehensive parts list of functional elements in human genome. This includes elements that act at protein and RNA levels, and regulatory elements that control cells and circumstances in which gene is active. Data from 2012-present.
View all literature mentionsDatabase for visualizing and making use of public ChIP-seq data. ChIP-Atlas covers almost all public ChIP-seq experiments and data submitted to the SRA (Sequence Read Archives) in NCBI, DDBJ, or ENA.
View all literature mentionsSoftware package for differential gene expression analysis based on the negative binomial distribution. Used for analyzing RNA-seq data for differential analysis of count data, using shrinkage estimation for dispersions and fold changes to improve stability and interpretability of estimates.
View all literature mentionsCell line GM12878 is a Transformed cell line with a species of origin Homo sapiens (Human)
View all literature mentionsCell line Jurkat is a Cancer cell line with a species of origin Homo sapiens (Human)
View all literature mentionsCell line HEK293T is a Transformed cell line with a species of origin Homo sapiens (Human)
View all literature mentions