Back

Robust and Adaptive Deep Model Ensemble Framework Fine-tuned by Structural Information for Drug-Target Interactions

Wei, J.; Zhuo, L.; Fu, X.; Zhang, J.; Zeng, X.; Zou, Q.

2023-10-23 bioinformatics
10.1101/2023.10.20.563031 bioRxiv
Show abstract

In the fields of new drug development and drug repositioning, drug-target interactions (DTI) play a pivotal role. Although deep learning models have already made significant contributions in this domain, the state-of-the-art models still exhibit shortcomings in predictive performance and issues of false-negative errors. Based on these observations, we constructed a streamlined yet effective base learner model. With our designed adaptive feature weight network, the model can capture key features within drugs (targets). Furthermore, by cross-partitioning the training data, multiple base learners are integrated into a powerful ensemble model named EADTN. The performance of the model is further enhanced as the number of base learners increases. Additionally, we employed a single-linkage clustering algorithm to cluster drugs and proteins and leveraged this clustering information to fine-tune the base learners, which elevates the value of EADTN in real-world applications like drug repositioning and targeted drug development. Our designed substructure importance ranking method also demonstrates the models exceptional capability to recognize key substructures. Benefiting from the models low generalization error capability, we successfully identified false-negative samples within the dataset, revealing new interaction relationships. Experimental results indicate that EADTN consistently outperforms existing state-of-the-art models across multiple datasets. More importantly, the ensemble learning and clustering fine-tuning approaches adopted by our model offer a fresh perspective for related fields.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.