Enhancing stochastic resonance using a reinforcement-learning based method. (April 2023)
- Record Type:
- Journal Article
- Title:
- Enhancing stochastic resonance using a reinforcement-learning based method. (April 2023)
- Main Title:
- Enhancing stochastic resonance using a reinforcement-learning based method
- Authors:
- Ding, Jianpeng
Lei, Youming - Abstract:
- We propose a new method to enhance stochastic resonance based on reinforcement learning, which does not require a priori knowledge of the underlying dynamics. The reward function of the reinforcement learning algorithm is determined by introducing a moving signal-to-noise ratio, which promptly quantifies the ratio of signal power to noise power by updating time series with a fixed length. To maximize the cumulative reward, the reward function can guide the actions to enhance the signal-to-noise ratio of systems as largely as possible with the help of the moving signal-to-noise ratio. Since the occurrence of the spike of excitable systems, which requires the systems to evolve for some time, should be considered an important component for the definition of the signal-to-noise ratio, the reward corresponding to the current moment cannot be obtained immediately and this usually results in a delayed reward. The delayed reward may cause the policy of the reinforcement learning algorithm to update with an incompatible reward, which affects the stability and convergence of the algorithm. To overcome this challenge, we devise a technique of double Q-tables, where one Q-table is used to generate actions, and the other is used to correct deviations. In this way, the policy can be updated with a corresponding reward, which ameliorates the stability of the algorithm and accelerates its convergence speed. We show with two illustrative examples, the Fitzhugh–Nagumo and Hindmarsh–RoseWe propose a new method to enhance stochastic resonance based on reinforcement learning, which does not require a priori knowledge of the underlying dynamics. The reward function of the reinforcement learning algorithm is determined by introducing a moving signal-to-noise ratio, which promptly quantifies the ratio of signal power to noise power by updating time series with a fixed length. To maximize the cumulative reward, the reward function can guide the actions to enhance the signal-to-noise ratio of systems as largely as possible with the help of the moving signal-to-noise ratio. Since the occurrence of the spike of excitable systems, which requires the systems to evolve for some time, should be considered an important component for the definition of the signal-to-noise ratio, the reward corresponding to the current moment cannot be obtained immediately and this usually results in a delayed reward. The delayed reward may cause the policy of the reinforcement learning algorithm to update with an incompatible reward, which affects the stability and convergence of the algorithm. To overcome this challenge, we devise a technique of double Q-tables, where one Q-table is used to generate actions, and the other is used to correct deviations. In this way, the policy can be updated with a corresponding reward, which ameliorates the stability of the algorithm and accelerates its convergence speed. We show with two illustrative examples, the Fitzhugh–Nagumo and Hindmarsh–Rose models, stochastic resonance is significantly enhanced by the proposed method for two typical types of stochastic resonances, classical stochastic resonance with a weak signal and coherent resonance without weak signals, respectively. We also show the robustness of the proposed method. … (more)
- Is Part Of:
- Journal of vibration and control. Volume 29:Number 7/8(2023)
- Journal:
- Journal of vibration and control
- Issue:
- Volume 29:Number 7/8(2023)
- Issue Display:
- Volume 29, Issue 7/8 (2023)
- Year:
- 2023
- Volume:
- 29
- Issue:
- 7/8
- Issue Sort Value:
- 2023-0029-NaN-0000
- Page Start:
- 1461
- Page End:
- 1471
- Publication Date:
- 2023-04
- Subjects:
- stochastic resonance -- reinforcement learning -- moving signal-to-noise ratio -- delayed reward -- signal processing
Vibration -- Periodicals
Damping (Mechanics) -- Periodicals
620.3 - Journal URLs:
- http://jvc.sagepub.com ↗
http://www.ingenta.com/journals/browse/sage/j324?mode=direct ↗
http://www.uk.sagepub.com/home.nav ↗
http://firstsearch.oclc.org ↗ - DOI:
- 10.1177/10775463211068895 ↗
- Languages:
- English
- ISSNs:
- 1077-5463
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 26003.xml