LG AI DB MLNov 29, 2018

Molecular Sets (MOSES): A Benchmarking Platform for Molecular Generation Models

Daniil Polykovskiy, Alexander Zhebrak, Benjamin Sanchez-Lengeling, Sergey Golovanov, Oktai Tatanov, Stanislav Belyaev, Rauf Kurbanov, Aleksey Artamonov, Vladimir Aladinskiy, Mark Veselov, Artur Kadurin, Simon Johansson

arXiv:1811.12823v531.9919 citationsHas Code

Originality Synthesis-oriented

AI Analysis

This work addresses the problem of inconsistent evaluation for researchers in generative chemistry, though it is incremental as it builds on existing models and datasets.

The authors tackled the lack of standardized comparison for molecular generative models by introducing the Molecular Sets (MOSES) benchmarking platform, which includes datasets and metrics, and they implemented and compared several models to provide reference points.

Generative models are becoming a tool of choice for exploring the molecular space. These models learn on a large training dataset and produce novel molecular structures with similar properties. Generated structures can be utilized for virtual screening or training semi-supervised predictive models in the downstream tasks. While there are plenty of generative models, it is unclear how to compare and rank them. In this work, we introduce a benchmarking platform called Molecular Sets (MOSES) to standardize training and comparison of molecular generative models. MOSES provides a training and testing datasets, and a set of metrics to evaluate the quality and diversity of generated structures. We have implemented and compared several molecular generation models and suggest to use our results as reference points for further advancements in generative chemistry research. The platform and source code are available at https://github.com/molecularsets/moses.

View on arXiv PDF Code

Similar