Search papers, labs, and topics across Lattice.
This study introduces a novel transition state (TS) search model leveraging the soft actor-critic reinforcement learning algorithm to navigate potential energy surfaces based on local energetic and curvature information. By framing the TS search as a sequential decision-making process, the agent autonomously proposes chemically meaningful structural updates without relying on labeled trajectories or predefined pathways. The method successfully identifies TS geometries for benchmark reactions, demonstrating a significant reduction in dependence on initial guesses and paving the way for scalable automated reaction discovery in chemistry.
Reinforcement learning can autonomously uncover transition states without the need for initial guesses or extensive training datasets, revolutionizing chemical reaction discovery.
Transition state (TS) search is a crucial step in understanding chemical reactivity and mechanisms, yet conventional algorithms remain computationally intensive and heavily reliant on initial guesses, user s expertise, and chemical intuition. While recent machine learning approaches have shown promise, they demand either large training datasets or geometric interpolation between known endpoints, limiting their generality. In this work, we introduce a TS search model based on the soft actor-critic model, an advanced reinforcement learning algorithm in which an agent learns to navigate potential energy surfaces directly from local energetic and curvature information starting from a given reactant and its corresponding product. By formulating the search as a sequential decision-making process in internal coordinates, the agent adaptively proposes chemically meaningful structural updates through a reward function designed to promote movement towards saddle point regions. Without labelled trajectories or prescribed reaction pathways, the method successfully identifies TS geometries for standard benchmark reactions, operating directly on realistic molecular potential energy surfaces. These results highlight the potential of RL as a general strategy for reducing dependence on initial guesses and enabling scalable, automated reaction discovery across diverse chemical systems.