Unknown

Dataset Information

0

Predicting retrosynthetic pathways using transformer-based models and a hyper-graph exploration strategy.


ABSTRACT: We present an extension of our Molecular Transformer model combined with a hyper-graph exploration strategy for automatic retrosynthesis route planning without human intervention. The single-step retrosynthetic model sets a new state of the art for predicting reactants as well as reagents, solvents and catalysts for each retrosynthetic step. We introduce four metrics (coverage, class diversity, round-trip accuracy and Jensen-Shannon divergence) to evaluate the single-step retrosynthetic models, using the forward prediction and a reaction classification model always based on the transformer architecture. The hypergraph is constructed on the fly, and the nodes are filtered and further expanded based on a Bayesian-like probability. We critically assessed the end-to-end framework with several retrosynthesis examples from literature and academic exams. Overall, the frameworks have an excellent performance with few weaknesses related to the training data. The use of the introduced metrics opens up the possibility to optimize entire retrosynthetic frameworks by focusing on the performance of the single-step model only.

SUBMITTER: Schwaller P 

PROVIDER: S-EPMC8152799 | biostudies-literature | 2020 Mar

REPOSITORIES: biostudies-literature

altmetric image

Publications

Predicting retrosynthetic pathways using transformer-based models and a hyper-graph exploration strategy.

Schwaller Philippe P   Petraglia Riccardo R   Zullo Valerio V   Nair Vishnu H VH   Haeuselmann Rico Andreas RA   Pisoni Riccardo R   Bekas Costas C   Iuliano Anna A   Laino Teodoro T  

Chemical science 20200303 12


We present an extension of our Molecular Transformer model combined with a hyper-graph exploration strategy for automatic retrosynthesis route planning without human intervention. The single-step retrosynthetic model sets a new state of the art for predicting reactants as well as reagents, solvents and catalysts for each retrosynthetic step. We introduce four metrics (coverage, class diversity, round-trip accuracy and Jensen-Shannon divergence) to evaluate the single-step retrosynthetic models,  ...[more]

Similar Datasets

2024-09-13 | GSE262953 | GEO
| S-EPMC8152431 | biostudies-literature
| S-EPMC5658761 | biostudies-literature
| PRJNA1094989 | ENA
| S-EPMC9944243 | biostudies-literature
| S-EPMC11302905 | biostudies-literature
| S-EPMC11345417 | biostudies-literature
| S-EPMC10191610 | biostudies-literature
| S-EPMC5530534 | biostudies-other
| S-EPMC11447720 | biostudies-literature