Directed Chemical Evolution via Navigating Molecular Encoding Space
Wang, L.; Wu, Y.; Luo, H.; Liang, M.; Zhou, Y.; Chen, C.; Liu, C.; Zhang, J.; Zhang, Y.
Show abstract
Deep-learning techniques have significantly advanced small-molecule drug discovery. However, a critical gap remains between representation learning and small molecule generations, limiting their effectiveness in developing new drugs. We introduce Ouroboros, a unified framework that integrates molecular representation learning with generative modeling, enabling efficient chemical space exploration using pre-trained molecular encodings. By reframing molecular generation as a process of encoding space compression and decompression, Ouroboros resolves the challenges associated with iterative molecular optimization and facilitates directed chemical evolution within the encoding space. Comprehensive experimental tests demonstrate that Ouroboros significantly outperforms conventional approaches across multiple drug discovery tasks, including ligand-based virtual screening, chemical property prediction, and multi-target inhibitor design and optimization. Unlike task-specific models in traditional approaches, Ouroboros leverages a unified framework to achieve superior performance across diverse applications. Ouroboros offers a novel and highly scalable protocol for rapid chemical space exploration, fostering a potential paradigm shift in AI-driven drug discovery.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Efficient Generation of Protein Pockets with PocketGen 97%
- PSICHIC: physicochemical graph neural network for learning protein-ligand interaction fingerprints from sequence data 94%
- TrustAffinity: accurate, reliable and scalable out-of-distribution protein-ligand binding affinity prediction using trustworthy deep learning 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.