Back

Interdisciplinary Study on Drug-Induced-Phospholipidosis of Repurposing Libraries through Machine Learning and Experimental Evaluation in Different Cell Lines

Kuzikov, M.; Kalman, A.; Reinshagen, J.; Huchting, J.; Qian, K.; Axelsson, H.; Oestling, P.; Seashore-Ludlow, B.; Gadiya, Y.; Gribbon, P.; Zaliani, A.

2025-07-09 molecular biology
10.1101/2025.07.08.663638 bioRxiv
Show abstract

Bigger PicturePhospholipidosis is a cellular condition characterized by the excessive accumulation of phospholipids within cells, that also can be induced by medications--especially those known as cationic amphiphilic drugs. While this phenomenon is not always harmful in itself, it can interfere with how cells respond in laboratory experiments and complicate the interpretation of drug screening results, leading to potential delays or failures in the development of new therapies. As the pharmaceutical industry increasingly relies on high-throughput screening and machine learning to accelerate early-stage drug discovery, there is growing interest in identifying compounds that may cause phospholipidosis before they move too far through the development pipeline. In response to this challenge, we at the EU-Horizon HLTH Remedi4ALL project compiled and analyzed one of the most comprehensive datasets to date--over 5,000 repurposed drugs tested across multiple cell lines--to better understand which compounds are likely to induce phospholipidosis. Using this dataset, we developed a machine learning model capable of predicting the risk of phospholipidosis based on a drugs chemical structure. By validating our approach across diverse experimental settings, our goal is to provide the broader scientific and pharmaceutical communities with a practical tool to flag problematic compounds early, ultimately helping bring safer, more effective treatments to patients faster. Highlights- A comprehensive dataset for a repurposing collection, comprising 5,228 compounds, for induction of phospholipidosis in A549-ACE2 and Vero-E6 cells -A machine learning model developed using in-house experimental phospholipidosis data across multiple cell lines -Exploration of chemical features that induce phospholipidosis and investigation of the relationships underlying different cell line susceptibilities -A case study: correlation of drug-induced phospholipidosis with antiviral compound effects against SARS-CoV-2 using phenotypic screening data Phospholipidosis (PLD), a cellular adverse effect that is, among others, caused by numerous cationic amphiphilic drugs. Interest is raised within pharma discovery to predict this phenomenon, as it can impact the outcome of phenotypic cellular screens and significantly delay drug development processes. The development of accurate and validated machine learning models for predicting drug-induced PLD across different cell lines and research centers could provide a valuable early application tool for the pharmaceutical industry, potentially accelerating drug discovery and reducing the risk of late-stage failures. We report here the assembly, curation, testing and modeling of one of the largest datasets of repurposed drugs (5K+) tested for PLD induction on different cell lines. A machine-learning classification method was developed and validated to predict whether molecules are prone to induce PLD effects when applied in cell-based screens. O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=187 SRC="FIGDIR/small/663638v1_ufig1.gif" ALT="Figure 1"> View larger version (67K): org.highwire.dtl.DTLVardef@1c4760org.highwire.dtl.DTLVardef@92084dorg.highwire.dtl.DTLVardef@15effbaorg.highwire.dtl.DTLVardef@1e73f4d_HPS_FORMAT_FIGEXP M_FIG O_FLOATNOGraphical abstractC_FLOATNO C_FIG

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
Journal of Chemical Information and Modeling
238 papers in training set
Top 0.3%
15.2%
2
ChemMedChem
16 papers in training set
Top 0.1%
10.7%
3
SLAS Discovery
25 papers in training set
Top 0.1%
10.7%
4
Scientific Reports
3612 papers in training set
Top 13%
6.3%
5
International Journal of Molecular Sciences
494 papers in training set
Top 1%
5.6%
6
Journal of Medicinal Chemistry
77 papers in training set
Top 0.3%
3.5%
50% of probability mass above
7
ACS Chemical Biology
167 papers in training set
Top 0.8%
3.2%
8
Chemical Science
73 papers in training set
Top 0.6%
2.4%
9
Computational and Structural Biotechnology Journal
242 papers in training set
Top 3%
2.1%
10
Communications Chemistry
48 papers in training set
Top 0.4%
2.1%
11
iScience
1154 papers in training set
Top 13%
2.1%
12
PLOS ONE
5266 papers in training set
Top 46%
1.9%
13
ACS Omega
105 papers in training set
Top 1%
1.9%
14
ACS Pharmacology & Translational Science
40 papers in training set
Top 0.3%
1.7%
15
Frontiers in Pharmacology
111 papers in training set
Top 1%
1.7%
16
Communications Biology
993 papers in training set
Top 13%
1.7%
17
npj Antimicrobials and Resistance
11 papers in training set
Top 0.1%
1.7%
18
Nature Communications
5641 papers in training set
Top 48%
1.4%
19
eLife
5828 papers in training set
Top 57%
1.1%
20
Journal of Cheminformatics
29 papers in training set
Top 0.7%
0.9%
21
Biomolecules
100 papers in training set
Top 2%
0.9%
22
PLOS Computational Biology
1863 papers in training set
Top 20%
0.9%
23
ACS Central Science
71 papers in training set
Top 1%
0.9%
24
Molecules
39 papers in training set
Top 1%
0.9%
25
Acta Pharmaceutica Sinica B
11 papers in training set
Top 0.3%
0.9%
26
Science Advances
1243 papers in training set
Top 30%
0.9%
27
Pharmaceuticals
34 papers in training set
Top 1%
0.9%
28
European Journal of Medicinal Chemistry
17 papers in training set
Top 0.2%
0.9%
29
Scientific Data
209 papers in training set
Top 3%
0.6%
30
Database
61 papers in training set
Top 1%
0.6%