Searching the Resource Information Network

Our searching services are busy right now. Please try again later

  • Register
X
Forgot Password

If you have forgotten your password you can enter your email here and get a temporary password sent to your email.

X

Leaving Community

Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.

No
Yes

Comprehensive characterization of the alternative splicing landscape in head and neck squamous cell carcinoma reveals novel events associated with tumorigenesis and the immune microenvironment.

Zhi-Xuan Li | Zi-Qi Zheng | Zhuo-Hui Wei | Lu-Lu Zhang | Feng Li | Li Lin | Rui-Qi Liu | Xiao-Dan Huang | Jia-Wei Lv | Fo-Ping Chen | Xiao-Jun He | Jia-Li Guan | Jia Kou | Jun Ma | Guan-Qun Zhou | Ying Sun
Theranostics | 2019

Alternative splicing (AS) has emerged as a key event in tumor development and microenvironment formation. However, comprehensive analysis of AS and its clinical significance in head and neck squamous cell carcinoma (HNSC) is urgently required. Methods: Genome-wide profiling of AS events using RNA-Seq data from The Cancer Genome Atlas (TCGA) program was performed in a cohort of 464 patients with HNSC. Cancer-associated AS events (CASEs) were identified between paired HNSC and adjacent normal tissues and evaluated in functional enrichment analysis. Splicing networks and prognostic models were constructed using bioinformatics tools. Unsupervised clustering of the CASEs identified was conducted and associations with clinical, molecular and immune features were analyzed. Results: We detected a total of 32,309 AS events and identified 473 CASEs in HNSC; among these, 91 were validated in an independent cohort (n = 15). Functional protein domains were frequently altered, especially by CASEs affecting cancer drivers, such as PCSK5. CASE parent genes were significantly enriched in pathways related to HNSC and the tumor immune microenvironment, such as the viral carcinogenesis (FDR < 0.001), Human Papillomavirus infection (FDR < 0.001), chemokine (FDR < 0.001) and T cell receptor (FDR < 0.001) signaling pathways. CASEs enriched in immune-related pathways were closely associated with immune cell infiltration and cytolytic activity. AS regulatory networks suggested a significant association between splicing factor (SF) expression and CASEs and might be regulated by SF methylation. Eighteen CASEs were identified as independent prognostic factors for overall and disease-free survival. Unsupervised clustering analysis revealed distinct correlations between AS-based clusters and prognosis, molecular characteristics and immune features. Immunogenic features and immune subgroups cooperatively depict the immune features of AS-based clusters. Conclusion: This comprehensive genome-wide analysis of the AS landscape in HNSC revealed novel AS events related to carcinogenesis and immune microenvironment, with implications for prognosis and therapeutic responses.

Pubmed ID: 31695792

Associated grants

None

Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.

This is a list of tools and resources that we have found mentioned in this publication.


circlize (tool)

RRID:SCR_002141

Software package that implements and enhances circular visualization in R. Due to natural born feature of R to draw statistical graphics, this package can provide more general and flexible way to visualize huge information in circular style.

View all literature mentions

Cytoscape (tool)

RRID:SCR_003032

Software platform for complex network analysis and visualization. Used for visualization of molecular interaction networks and biological pathways and integrating these networks with annotations, gene expression profiles and other state data.

View all literature mentions

Gene Set Enrichment Analysis (tool)

RRID:SCR_003199

Software package for interpreting gene expression data. Used for interpretation of a large-scale experiment by identifying pathways and processes.

View all literature mentions

PROSITE (tool)

RRID:SCR_003457

Database of protein families and domains that is based on the observation that, while there is a huge number of different proteins, most of them can be grouped, on the basis of similarities in their sequences, into a limited number of families. Proteins or protein domains belonging to a particular family generally share functional attributes and are derived from a common ancestor. It is complemented by ProRule, a collection of rules based on profiles and patterns, which increases the discriminatory power of profiles and patterns by providing additional information about functionally and/or structurally critical amino acids. ScanProsite finds matches of your protein sequences to PROSITE signatures. PROSITE currently contains patterns and profiles specific for more than a thousand protein families or domains. Each of these signatures comes with documentation providing background information on the structure and function of these proteins. The database is available via FTP.

View all literature mentions

Pfam (tool)

RRID:SCR_004726

A database of protein families, each represented by multiple sequence alignments and hidden Markov models (HMMs). Users can analyze protein sequences for Pfam matches, view Pfam family annotation and alignments, see groups of related families, look at the domain organization of a protein sequence, find the domains on a PDB structure, and query Pfam by keywords. There are two components to Pfam: Pfam-A and Pfam-B. Pfam-A entries are high quality, manually curated families that may automatically generate a supplement using the ADDA database. These automatically generated entries are called Pfam-B. Although of lower quality, Pfam-B families can be useful for identifying functionally conserved regions when no Pfam-A entries are found. Pfam also generates higher-level groupings of related families, known as clans (collections of Pfam-A entries which are related by similarity of sequence, structure or profile-HMM).

View all literature mentions

STRING (tool)

RRID:SCR_005223

Database of known and predicted protein interactions. The interactions include direct (physical) and indirect (functional) associations and are derived from four sources: Genomic Context, High-throughput experiments, (Conserved) Coexpression, and previous knowledge. STRING quantitatively integrates interaction data from these sources for a large number of organisms, and transfers information between these organisms where applicable. The database currently covers 5''214''234 proteins from 1133 organisms. (2013)

View all literature mentions

SpliceSeq (tool)

RRID:SCR_005267

A Java application to investigate alternative mRNA splicing patterns in data from high-throughput mRNA sequencing studies. Sequence reads are mapped to splice graphs that unambiguously quantify the inclusion level of each exon and splice junction. The graphs are then traversed to predict the protein isoforms that are likely to result from the observed exon and splice junction reads. UniProt annotations are mapped to each protein isoform to identify potential functional impacts of alternative splicing. This tool may be used on a single RNASeq sample to identify genes with multiple spliceforms, on a pair of samples to identify differential splicing between the two, or on groups of samples to identify statistically significant group level differences in splicing patterns. SpliceSeq can be run from the install page as a java web start application to explore the sequencing data on their server or can be installed locally to analyze your own mRNA-Seq data.

View all literature mentions

Circos (tool)

RRID:SCR_011798

A software package for visualizing data and information. It visualizes data in a circular layout - this makes Circos ideal for exploring relationships between objects or positions.

View all literature mentions

KEGG (tool)

RRID:SCR_012773

Integrated database resource consisting of 16 main databases, broadly categorized into systems information, genomic information, and chemical information. In particular, gene catalogs in completely sequenced genomes are linked to higher-level systemic functions of cell, organism, and ecosystem. Analysis tools are also available. KEGG may be used as reference knowledge base for biological interpretation of large-scale datasets generated by sequencing and other high-throughput experimental technologies.

View all literature mentions

cBioPortal (tool)

RRID:SCR_014555

A portal that provides visualization, analysis and download of large-scale cancer genomics data sets.

View all literature mentions