Sparse Representation for High-dimensional Multiclass Microarray Data Classification
Miri, M.; Sadeghi, M. T.; abootalebi, v.
Show abstract
Sparse representation of signals has achieved satisfactory results in classification applications compared to the conventional methods. Microarray data, which are obtained from monitoring the expression levels of thousands of genes simultaneously, have very high dimensions in relation to the small number of samples. This has led to the weaknesses of state-of-the-art classifiers to cope with the microarray data classification problem. The ability of the sparse representation to represent the signals as a linear combination of a small number of training data and to provide a brief description of signals led to reducing computational complexity as well as increasing classification accuracy in many applications. Using all training samples in the dictionary imposes a high computational burden on the sparse coding stage of high dimensional data. Proposed solutions to solve this problem can be roughly divided into two categories: selection of a subset of training data using different criteria, or learning a concise dictionary. Another important factor in increasing the speed and accuracy of a sparse representation-based classifier is the algorithm which is used to solve the related {ell}1 -norm minimization problem. In this paper, different sparse representation-based classification methods are investigated in order to tackle the problem of 14-Tumors microarray data classification. Our experimental results show that good performances are obtained by selecting a subset of the original atoms and learning the associated dictionary. Also, using SL0 sparse coding algorithm increases speed, and in most cases, accuracy of the classifiers.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Multi- Stage Feature Selection (MSFS) Algorithm for UWB- Based Early Breast Cancer Size Prediction 96%
- Cardiac disease diagnosis based on GAN in case of missing data 96%
- Projection in genomic analysis: A theoretical basis to rationalize tensor decomposition and principal component analysis as feature selection tools 96%
Similar papers in this journal
- Hilbert-Envelope Features for Cardiac Disease Classification from Noisy Phonocardiograms 94%
- Evaluating three different adaptive decomposition methods for EEG signal seizure detection and classification 94%
- Improved online event detection and differentiation by a simple gradient-based nonlinear transformation: Implications for the biomedical signal and image analysis 94%
Similar papers in this journal
Similar papers in this journal
- Decoding Clinical Biomarker Space of COVID-19: Exploring Matrix Factorization-based Feature Selection Methods 97%
- A machine-learning Approach for Stress Detection Using Wearable Sensors in Free-living Environments 94%
- Radiomics Analysis Using Stability Selection Supervised Principal Component Analysis for Right-censored Survival Data 93%
Similar papers in this journal
- Estimation of Three-Dimensional Chromatin Morphology for Nuclear Classification and Characterisation 96%
- A Robust Spike Sorting Method based on the Joint Optimization of Linear Discrimination Analysis and Density Peaks 95%
- A Convolution Based Computational Approach Towards DNA N6-methyladenine Site Identification and Motif Extraction in Rice Genome 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.