G2GT: Retrosynthesis Prediction with Graph to Graph Attention Neural Network and Self-Training

Zaiyun Lin, Shiqiu Yin, Lei Shi, Wenbiao Zhou, YingSheng Zhang

2022-04-19Ensemble Learning Data Augmentation Retrosynthesis Single-step retrosynthesis Graph Attention

Abstract

Retrosynthesis prediction is one of the fundamental challenges in organic chemistry and related fields. The goal is to find reactants molecules that can synthesize product molecules. To solve this task, we propose a new graph-to-graph transformation model, G2GT, in which the graph encoder and graph decoder are built upon the standard transformer structure. We also show that self-training, a powerful data augmentation method that utilizes unlabeled molecule data, can significantly improve the model's performance. Inspired by the reaction type label and ensemble learning, we proposed a novel weak ensemble method to enhance diversity. We combined beam search, nucleus, and top-k sampling methods to further improve inference diversity and proposed a simple ranking algorithm to retrieve the final top-10 results. We achieved new state-of-the-art results on both the USPTO-50K dataset, with top1 accuracy of 54%, and the larger data set USPTO-full, with top1 accuracy of 50%, and competitive top-10 results.

Results

Task	Dataset	Metric	Value	Model
Single-step retrosynthesis	USPTO-50k	Top-1 accuracy	54.1	G2GT (reaction class unknown)
Single-step retrosynthesis	USPTO-50k	Top-10 accuracy	77.7	G2GT (reaction class unknown)
Single-step retrosynthesis	USPTO-50k	Top-3 accuracy	69.9	G2GT (reaction class unknown)
Single-step retrosynthesis	USPTO-50k	Top-5 accuracy	74.5	G2GT (reaction class unknown)

Related Papers

Simulate, Refocus and Ensemble: An Attention-Refocusing Scheme for Domain Generalization2025-07-17 Overview of the TalentCLEF 2025: Skill and Job Title Intelligence for Human Capital Management2025-07-17 Pixel Perfect MegaMed: A Megapixel-Scale Vision-Language Foundation Model for Generating High Resolution Medical Images2025-07-17 Similarity-Guided Diffusion for Contrastive Sequential Recommendation2025-07-16 Catching Bid-rigging Cartels with Graph Attention Neural Networks2025-07-16 Data Augmentation in Time Series Forecasting through Inverted Framework2025-07-15 Iceberg: Enhancing HLS Modeling with Synthetic Data2025-07-14 Wavelet-Enhanced Neural ODE and Graph Attention for Interpretable Energy Forecasting2025-07-14