COMBINING CORRELATION-BASED AND REWARD-BASED LEARNING IN NEURAL CONTROL FOR POLICY IMPROVEMENT-Reference-Cited by-同舟云学术

COMBINING CORRELATION-BASED AND REWARD-BASED LEARNING IN NEURAL CONTROL FOR POLICY IMPROVEMENT

Published:2013-05 Issue:02n03 Volume:16 Page:1350015
ISSN:0219-5259
Container-title:Advances in Complex Systems
language:en
Short-container-title:Advs. Complex Syst.

Author:

MANOONPONG PORAMATE¹²,KOLODZIEJSKI CHRISTOPH¹,WÖRGÖTTER FLORENTIN¹,MORIMOTO JUN¹²

Affiliation:

1. Bernstein Center for Computational Neuroscience, The Third Institute of Physics, University of Göttingen, Göttingen 37077, Germany

2. ATR Computational Neuroscience Laboratories, 2-2-2 Hikaridai Seika-cho, Soraku-gun, Kyoto 619-0288, Japan

Abstract

Classical conditioning (conventionally modeled as correlation-based learning) and operant conditioning (conventionally modeled as reinforcement learning or reward-based learning) have been found in biological systems. Evidence shows that these two mechanisms strongly involve learning about associations. Based on these biological findings, we propose a new learning model to achieve successful control policies for artificial systems. This model combines correlation-based learning using input correlation learning (ICO learning) and reward-based learning using continuous actor–critic reinforcement learning (RL), thereby working as a dual learner system. The model performance is evaluated by simulations of a cart-pole system as a dynamic motion control problem and a mobile robot system as a goal-directed behavior control problem. Results show that the model can strongly improve pole balancing control policy, i.e., it allows the controller to learn stabilizing the pole in the largest domain of initial conditions compared to the results obtained when using a single learning mechanism. This model can also find a successful control policy for goal-directed behavior, i.e., the robot can effectively learn to approach a given goal compared to its individual components. Thus, the study pursued here sharpens our understanding of how two different learning mechanisms can be combined and complement each other for solving complex tasks.

Publisher

World Scientific Pub Co Pte Lt

Subject

Control and Systems Engineering

Link

https://www.worldscientific.com/doi/pdf/10.1142/S021952591350015X

Reference46 articles.

1. ACTION DISCOVERY FOR SINGLE AND MULTI-AGENT REINFORCEMENT LEARNING

2. Neuronlike adaptive elements that can solve difficult learning control problems

Cited by 10 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. A framework for visual-based adaptive object-robot interaction of a mobile service robot;Adaptive Behavior;2024-04-27

2. NeuroVis: Real-Time Neural Information Measurement and Visualization of Embodied Neural Systems;Frontiers in Neural Circuits;2021-12-27

3. Locomotion Control With Frequency and Motor Pattern Adaptations;Frontiers in Neural Circuits;2021-11-25

4. A neuroplasticity-inspired neural circuit for acoustic navigation with obstacle avoidance that learns smooth motion paths;Neural Computing and Applications;2018-11-08

5. An Adaptive Neural Mechanism for Acoustic Motion Perception with Varying Sparsity;Frontiers in Neurorobotics;2017-03-09