Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
Long noncoding RNAs (lncRNAs) are regulators of cell differentiation and development. The lncRNA transcriptome in human hematopoietic stem and progenitor cells is not comprehensively defined. We investigated lncRNAs in 979 human bone marrow-derived CD34+ cells by single cell RNA sequencing followed by de novo transcriptome reconstruction. We identified 3,173 lncRNAs in total, among which 2,365 were previously unknown, and we characterized lncRNA stem, differentiation, and maturation signatures. lncRNA expression exhibited high cell-to-cell variation, which was only apparent in single cell analysis. lncRNA expression followed a lineage-specific and highly dynamic pattern during early hematopoiesis. lncRNAs in hematopoietic cells closely correlated with protein-coding genes of known functions in the regulation of hematopoiesis and cell fate decisions, and the potential regulatory roles of lncRNAs in hematopoiesis were imputed by projection from protein-coding genes with a "guilt-by-association" approach. We characterized lncRNAs preferentially expressed in hematopoietic stem cells and in various downstream differentiated lineage progenitors. We also profiled lncRNA expression in single cells from patients with myelodysplastic syndromes and in aneuploid cells in particular. Our study provides a global view of lncRNAs in human hematopoietic stem and progenitor cells. We observed a highly ordered pattern of lncRNA expression and participation in regulation of early hematopoiesis, and coordinate aberrant messenger RNA and lncRNA transcriptomes in dysplastic hematopoiesis. (Registered at clinicaltrials.gov with identifiers: 00001620, 00001397).
Pubmed ID: 30545929
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Web application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.
View all literature mentionsRegistry and results database of federally and privately supported clinical trials conducted in United States and around world. Provides information about purpose of trial, who may participate, locations, and phone numbers for more details. This information should be used in conjunction with advice from health care professionals.Offers information for locating federally and privately supported clinical trials for wide range of diseases and conditions. Research study in human volunteers to answer specific health questions. Interventional trials determine whether experimental treatments or new ways of using known therapies are safe and effective under controlled environments. Observational trials address health issues in large groups of people or populations in natural settings. ClinicalTrials.gov contains trials sponsored by National Institutes of Health, other federal agencies, and private industry. Studies listed in database are conducted in all 50 States and in 178 countries.
View all literature mentionsA database of protein families, each represented by multiple sequence alignments and hidden Markov models (HMMs). Users can analyze protein sequences for Pfam matches, view Pfam family annotation and alignments, see groups of related families, look at the domain organization of a protein sequence, find the domains on a PDB structure, and query Pfam by keywords. There are two components to Pfam: Pfam-A and Pfam-B. Pfam-A entries are high quality, manually curated families that may automatically generate a supplement using the ADDA database. These automatically generated entries are called Pfam-B. Although of lower quality, Pfam-B families can be useful for identifying functionally conserved regions when no Pfam-A entries are found. Pfam also generates higher-level groupings of related families, known as clans (collections of Pfam-A entries which are related by similarity of sequence, structure or profile-HMM).
View all literature mentionsExpression profiling and promoter identification software tool for transcriptional network analysis and transcriptome characterization. DeepCAGE, the combination of next-generation sequencing with next generation expression profiling provides unsurpassed solutions for expression profiling and genome annotation. CAGE will be the experimental approach at need to link gene expression and control regions in the genome. With the availability of next-generation sequencing methods, DNAFORM now offers DeepCAGE services. DeepCAGE libraries are prepared for direct analysis by an Illumina/Solexa Sequencer. One sequencing run using one channel on an Illumina/Solexa Sequencer can yield in over 4,000,000 reads per sample. CAGE is based on our full-length cDNA library technology, where an adaptor is ligated to the 5''''-end of full-length cDNAs, which introduces a recognition site for a Class IIs restriction endonuclease adjacent to the 5''''-end of the cDNA. The Class IIs restriction endonuclease, here MmeI, allows for the cloning of short tags as derived from the 5''''-end of transcripts into concatemers for high-throughput sequencing. CAGE tags are further characterized by mapping to genomic sequences, which enables the identification of transcriptional start sites. As such CAGE can contribute to projects in Gene Discovery, Gene Expression, and Promoter Identification. After the genome sequencing projects have provided us with the genetic blueprints for many organisms, new questions have to be answered on how to correlate the observed genotypes with related phenotypes, and how to understand the regulation of genetic information in time and space. The dynamics of living systems and the functional behavior of cells in multicellular organisms has thus become the subject of the emerging field of system biology. Integration of experimental approaches and computer aided theories on a system level will be the fundamental principle to drive systems biology in order to understand the principles behind complex regulatory networks, which will be an ambitious goal requiring new approaches in life sciences. For ordering and additional information, please contact us under contact_at_dnaform.jp
View all literature mentionsSoftware analysis package for molecular biology community. Automatically copes with data in variety of formats and allows transparent retrieval of sequence data from web. Libraries are provided with package. Provides toolkit for creating bioinformatics applications or workflows. Provides set of sequence analysis programs. Provided programs cover areas such as sequence alignment, rapid database searching with sequence patterns, protein motif identification, nucleotide sequence pattern analysis, codon usage analysis for small genomes, rapid identification of sequence patterns in large scale sequence sets, and presentation tools for publication.
View all literature mentionsSoftware tool for transcriptome assembly and differential expression analysis for RNA-Seq. Includes script called cuffmerge that can be used to merge together several Cufflinks assemblies. It also handles running Cuffcompare as well as automatically filtering a number of transfrags that are likely to be artifacts. If the researcher has a reference GTF file, the researcher can provide it to the script to more effectively merge novel isoforms and maximize overall assembly quality.
View all literature mentions