Distance preserving dimension reduction withlocal-topology based scaling for improvedclassification of Biomedical data-sets.
Khosla, K.; Jha, I. P.; KUMAR, V.
Show abstract
Dimension reduction is often used for several procedures of analysis of high dimensional biomedical data-sets such as classification or outlier detection. To improve performance of such data-mining steps, preserving both distance information and local topology among data-points could be more useful than giving priority to visualisation in low dimension. Therefore, we introduce topology preserving distance scaling (TPDS) to augment dimension reduction method meant to reproduce distance information in higher dimension. Our approach involves distance inflation to preserve local topology to avoid collapse during distance preservation based optimisation. Applying TPDS on diverse biomedical data-sets revealed that besides providing better visualisation than typical distance preserving methods, TPDS leads to better classification of data points in reduced dimension. For data-sets with outliers, the approach of TPDS also proves to be useful, even for purely distance-preserving method for achieving better convergence.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Cardiac disease diagnosis based on GAN in case of missing data 95%
- Improving prediction of drug-target interactions based on fusing multiple features with data balancing and feature selection techniques 95%
- A Machine Learning Model of Microscopic Agglutination Test for Diagnosis of Leptospirosis 94%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- A Convolution Based Computational Approach Towards DNA N6-methyladenine Site Identification and Motif Extraction in Rice Genome 95%
- Comparing protein-protein interaction networks of SARS-CoV-2 and (H1N1) influenza using topological features 95%
- A Robust Spike Sorting Method based on the Joint Optimization of Linear Discrimination Analysis and Density Peaks 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.