Back

Mixture-of-Experts Approach for Enhanced Drug-Target Interaction Prediction and Confidence Assessment

Lu, Y.; Lee, S.; Kang, S.; Kim, S.

2024-08-08 bioinformatics
10.1101/2024.08.06.606753 bioRxiv
Show abstract

In recent years, numerous deep learning models have been developed for drug-target interaction (DTI) prediction. These DTI models specialize in handling data with distinct distributions and features, often yielding inconsistent predictions when applied to unseen data points. This inconsistency poses a challenge for researchers aiming to utilize these models in downstream drug development tasks. Particularly in screening potential active compounds, providing a ranked list of candidates that likely interact with the target protein can guide scientists in prioritizing their experimental efforts. However, achieving this is difficult as each current DTI model can provide a different list based on its learned feature space. To address these issues, we propose EnsDTI, a Mixture-of-Experts architecture designed to enhance the performance of existing DTI models for more reliable drug-target interaction predictions. We integrate an inductive conformal predictor to provide confidence scores for each prediction, enabling EnsDTI to offer a reliable list of candidates for a specific target. Empirical evaluations on four benchmark datasets demonstrate that EnsDTI not only improves DTI prediction performance with an average accuracy improvement of 2.7% compared to the best performing baseline, but also offers a reliable ranked list of candidate drugs with the highest confidence, showcasing its potential for ranking potential active compounds in future applications. CCS CONCEPTS* Applied computing [->] Bioinformatics; * Computing methodologies [->] Artificial intelligence.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.