Back

Benchmarking machine learning models for the analysis of genetic data using FRESA.CAD Binary Classification Benchmarking

De Velasco Oriol, J.; Martinez-Torteya, A.; Trevino, V.; Alanis, I.; Vallejo, E. E.; Tamez-Pena, J. G.

2019-08-13 bioinformatics
10.1101/733675 bioRxiv
Show abstract

BackgroundMachine learning models have proven to be useful tools for the analysis of genetic data. However, with the availability of a wide variety of such methods, model selection has become increasingly difficult, both from the human and computational perspective.\n\nResultsWe present the R package FRESA.CAD Binary Classification Benchmarking that performs systematic comparisons between a collection of representative machine learning methods for solving binary classification problems on genetic datasets.\n\nConclusionsFRESA.CAD Binary Benchmarking demonstrates to be a useful tool over a variety of binary classification problems comprising the analysis of genetic data showing both quantitative and qualitative advantages over similar packages.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.