Learning Chemical Reaction Representation with Reactant-Product Alignment

Zeng, Kaipeng; Liu, Xianbin; Zhang, Yu; Yang, Xiaokang; Jin, Yaohui; Xu, Yanyan

Computer Science > Machine Learning

arXiv:2411.17629 (cs)

[Submitted on 26 Nov 2024 (v1), last revised 3 Jan 2025 (this version, v2)]

Title:Learning Chemical Reaction Representation with Reactant-Product Alignment

Authors:Kaipeng Zeng, Xianbin Liu, Yu Zhang, Xiaokang Yang, Yaohui Jin, Yanyan Xu

View PDF HTML (experimental)

Abstract:Organic synthesis stands as a cornerstone of the chemical industry. The development of robust machine learning models to support tasks associated with organic reactions is of significant interest. However, current methods rely on hand-crafted features or direct adaptations of model architectures from other domains, which lack feasibility as data scales increase or ignore the rich chemical information inherent in reactions. To address these issues, this paper introduces RAlign, a novel chemical reaction representation learning model for various organic reaction-related tasks. By integrating atomic correspondence between reactants and products, our model discerns the molecular transformations that occur during the reaction, thereby enhancing comprehension of the reaction mechanism. We have designed an adapter structure to incorporate reaction conditions into the chemical reaction representation, allowing the model to handle various reaction conditions and to adapt to various datasets and downstream tasks. Additionally, we introduce a reaction-center-aware attention mechanism that enables the model to concentrate on key functional groups, thereby generating potent representations for chemical reactions. Our model has been evaluated on a range of downstream tasks. Experimental results indicate that our model markedly outperforms existing chemical reaction representation learning architectures on most of the datasets. We plan to open-source the code contingent upon the acceptance of the paper.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2411.17629 [cs.LG]
	(or arXiv:2411.17629v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2411.17629

Submission history

From: Kaipeng Zeng [view email]
[v1] Tue, 26 Nov 2024 17:41:44 UTC (1,606 KB)
[v2] Fri, 3 Jan 2025 16:55:38 UTC (4,896 KB)

Computer Science > Machine Learning

Title:Learning Chemical Reaction Representation with Reactant-Product Alignment

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Learning Chemical Reaction Representation with Reactant-Product Alignment

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators