vix.ing · top · new · best · stats · spec

Improving Targeted Molecule Generation through Language Model Fine-Tuning Via Reinforcement Learning

2024/05/10 by Salma J. Ahmed, Ahmed, Salma J., Mohammed, Emad A.
Biochemistry, Genetics and Molecular Biology · Materials Science · #Biomolecules (q-bio.BM) #Chemical Synthesis and Analysis #FOS: Biological sciences #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science

paper · pdf · doi:10.48550/arxiv.2405.06836

openalex publication_date 2024/05/10 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

Developing new drugs is laborious and costly, demanding extensive time investment. In this paper, we introduce a de-novo drug design strategy, which harnesses the capabilities of language models to devise targeted drugs for specific proteins. Employing a Reinforcement Learning (RL) framework utilizing Proximal Policy Optimization (PPO), we refine the model to acquire a policy for generating drugs tailored to protein targets. The proposed method integrates a composite reward function, combining considerations of drug-target interaction and molecular validity. Following RL fine-tuning, the proposed method demonstrates promising outcomes, yielding notable improvements in molecular validity, interaction efficacy, and critical chemical properties, achieving 65.37 for Quantitative Estimation of Drug-likeness (QED), 321.55 for Molecular Weight (MW), and 4.47 for Octanol-Water Partition Coefficient (logP), respectively. Furthermore, out of the generated drugs, only 0.041% do not exhibit novelty.

Related