Evolution Strategies as an Alternative to Reinforcement Learning in De Novo Molecular Generation - Evaluating Distribution Based Gaussian Evolution Strategies in REINVENT vs. Policy Based Reinforcement Learning on the Practical Molecular Optimization Benchmark
Hämtar...
Ladda ner
Publicerad
Författare
Typ
Examensarbete för masterexamen
Master's Thesis
Master's Thesis
Modellbyggare
Tidskriftstitel
ISSN
Volymtitel
Utgivare
Sammanfattning
De novo molecular generation examines how computational methods can propose
novel drug candidates by scoring generated molecules against desired properties and
updating generation toward higher scoring regions. One approach trains a SMILES
based language model and fine tunes it toward such regions. Fine tuning is commonly
performed with policy gradient reinforcement learning (RL), as in AstraZenecas
highly optimized REINVENT platform. An alternative is evolutionary strategies
(ES), where a population of models is created and their parameters are updated
to bias generation toward higher scores. This thesis investigates how distribution
based ES compares with RL when implemented in REINVENT and evaluated on the
Practical Molecular Optimization benchmark. OpenAI-ES shows near competitive
performance on scoring and diversity metrics, whereas variance estimating Natural
ES methods perform poorly. Variations of fixed variance ES are explored, and a novel
sampling technique that biases generation toward high diversity shows promising
performance across multiple metrics.
Beskrivning
Ämne/nyckelord
Evolution Strategies, Reinforcement Learning, De Novo Molecular Gen eration, Molecular AI, Drug Development, Fine-Tuning
