Robust Classification of Immune Subtypes in Cancer
Gibbs, D. L.
Show abstract
As part of the immune landscape of cancer, six immune subtypes were defined which describe a categorization of tumor-immune states. A number of phenotypic variables were found to associate with immune subtypes, such as nonsilent mutation rates, regulation of immunomodulator genes, and cytokine network structures. An ensemble classifier based on XGBoost is introduced with the goal of classifying tumor samples into one of six immune subtypes. Robust performance was accomplished through feature engineering; quartile-levels, binary gene-pair features, and gene-set-pair features were computed for each sample independently. The classifier is robust to software pipeline and normalization scheme, making it applicable to any expression data format from raw count data to TPMs since the classification is essentially based on simple binary gene-gene level comparisons within a given sample. The classifier is available as an R package or part of the CRI iAtlas portal. Code / Tool availabilitySource Code https://github.com/Gibbsdavidl/ImmuneSubtypeClassifier Web App Tool https://www.cri-iatlas.org/
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- pyCancerSig: subclassifying human cancer with comprehensive single nucleotide, structural and microsatellite mutational signature deconstruction from whole genome sequencing 92%
- driveR: A Novel Method for Prioritizing Cancer Driver Genes Using Somatic Genomics Data 92%
- GEOlimma: Differential Expression Analysis and Feature Selection Using Pre-Existing Microarray Data 92%
Similar papers in this journal
Similar papers in this journal
- Topological embedding and directional feature importance in ensemble classifiers for multi-class classification 93%
- Wide and Deep Learning for Automatic Cell Type Identification 93%
- Modeling and analysis of site-specific mutations in cancer identifies known plus putative novel hotspots and bias due to contextual sequences 92%
Similar papers in this journal
- Blood-based transcriptomic signature panel identification for cancer diagnosis: Benchmarking of feature extraction methods 92%
- An approach for normalization and quality control for NanoString RNA expression data 92%
- Hierarchical cell-type identifier accurately distinguishes immune-cell subtypes enabling precise profiling of tissue microenvironment with single-cell RNA-sequencing 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.