Outline

Ingegneria Sismica

Ingegneria Sismica

Reinforcement Learning-Based Dynamic Optimization of Reaction Conditions and Product Selectivity Control for Electrocatalytic CO2 Reduction

Author(s): Hongliang Dong1, Yaxin Geng1, Jingjing Wang1
1School of Chemical Engineering, Hebei University of Technology, Tianjin 300130, China.
Dong, Hongliang., Geng, Yaxin., and Wang, Jingjing. “Reinforcement Learning-Based Dynamic Optimization of Reaction Conditions and Product Selectivity Control for Electrocatalytic CO2 Reduction.” Ingegneria Sismica Volume 43 Issue 3: 1-21, doi:10.65102/is20261243.

Abstract

Electrocatalytic  reduction reaction  is one of the important technical paths to realize the carbon neutrality strategy, but the reaction faces core challenges such as high sensitivity to operating conditions, severe competitive hydrogen evolution reaction, and difficulty in precise regulation of product selectivity. Traditional catalyst design and static condition optimization methods cannot adapt to the requirements of dynamic working conditions, so there is an urgent need to develop new intelligent regulation strategies. This paper proposes a dynamic optimization framework for electrocatalytic  reduction reaction conditions and product selectivity regulation based on deep reinforcement learning. The  process is modeled as a Markov Decision Process (MDP), a digital twin simulation environment integrating density functional theory (DFT), microkinetic models and experimental data is constructed, a state space including real-time potential, current density,  and intermediate coverage is designed, as well as a continuous action space focusing on the adjustment of potential step size, electrolyte concentration and flow rate. A multi-objective weighted reward function that takes into account Faradaic efficiency, energy conversion efficiency and long-term stability is proposed, and the Proximal Policy Optimization (PPO) algorithm is adopted to realize online dynamic regulation. Verified by simulation training and real flow electrolyzer experiments, the results show that the reinforcement learning optimization strategy enables the Faradaic efficiency of CO to reach  , the Faradaic efficiency of  product to reach 68.7%, the energy conversion efficiency to be increased to 56.3%, and the long-term operation stability to exceed 120 h, which are 27.1, 27.2 and 17.6 percentage points higher than those under traditional fixed conditions respectively. Furthermore, dynamic and fast switching between CO and  products is realized (response time < 5 min, selectivity stabilized above 85%), and the microscopic mechanism of regulation is revealed through the analysis of the dynamic evolution of intermediate coverage. Multi-objective Pareto frontier analysis verifies the flexibility of the framework in the efficiency-selectivity trade-off. The work in this paper breaks through the limitations of traditional static optimization, provides a new method and new paradigm for the intelligent and real-time regulation of electrocatalytic  reduction reaction, and has important theoretical significance and engineering application value for promoting the efficient resource utilization of  under the background of carbon neutrality.

Keywords
reinforcement learning; electrocatalytic 〖CO〗_2 reduction; dynamic condition optimization; product selectivity regulation; digital twin; multi-objective optimization; carbon neutrality

Related Articles

Zhihao Jiang1,2, Limi Chen1,2, Jing Yang1
1Hainan Vocational University of Science and Technology, Haikou 571126, China
2Institute for Mathematical Research, Universiti Putra Malaysia, Serdang 43400, Malaysia
Limi Chen1,2, Zhihao Jiang1,2, Jing Yang1
1Hainan Vocational University of Science and Technology, Haikou 571126, China
2Institute for Mathematical Research, Universiti Putra Malaysia, Serdang 43400, Malaysia
Hui Yuan1, Minjie Chai2, Siqing Xu1, Jinsong Li1, Jinwan Zheng1
1Electric Power Research Institute, State Grid Shanxi Electric Power Co., Ltd., Taiyuan, 030001, Shanxi, China
2Jincheng Power Supply Branch, State Grid Shanxi Electric Power Co., Ltd., Jincheng, 048000, Shanxi, China
Yanhan Zhu1,2
1China Academy of Cultural Heritage, Chaoyang District, 100029, Beijing, China
2Beijing University of Civil Engineering and Architecture, Xicheng District, 100044, Beijing, China
Ken Wang1, Jinhan Shu2, Kan Yuan1
1School of Digital Media, Shenzhen Polytechnic University, Shenzhen 518055, Guangdong, China
2Postdoctoral Mobile Station of Journalism and communication, Fudan University, Shanghai 200433, Shanghai, China