A Method to Improve the Performance of Reinforcement Learning Based on the Y Operator for a Class of Stochastic Differential Equation-Based Child-Mother Systems
–arXiv.org Artificial Intelligence
This paper introduces a novel operator, termed the Y operator, to elevate control performance in Actor-Critic(AC) based reinforcement learning for systems governed by stochastic differential equations(SDEs). The Y operator ingeniously integrates the stochasticity of a class of child-mother system into the Critic network's loss function, yielding substantial advancements in the control performance of RL algorithms.Additionally, the Y operator elegantly reformulates the challenge of solving partial differential equations for the state-value function into a parallel problem for the drift and diffusion functions within the system's SDEs.A rigorous mathematical proof confirms the operator's validity.This transformation enables the Y Operator-based Reinforcement Learning(YORL) framework to efficiently tackle optimal control problems in both model-based and data-driven systems.The superiority of YORL is demonstrated through linear and nonlinear numerical examples showing its enhanced performance over existing methods post convergence.
arXiv.org Artificial Intelligence
Jan-1-2024
- Country:
- North America > United States
- Texas (0.04)
- Pennsylvania > Allegheny County
- Pittsburgh (0.04)
- Europe > Germany
- North Rhine-Westphalia > Cologne Region > Aachen (0.04)
- Asia
- Middle East > Jordan (0.04)
- China
- Hubei Province > Wuhan (0.04)
- Hong Kong (0.04)
- North America > United States
- Genre:
- Research Report (0.82)
- Industry:
- Energy (0.67)
- Automobiles & Trucks (0.46)
- Transportation > Ground
- Road (0.46)
- Technology: