Multiple Model-Based Reinforcement Learning

Top Cited Papers

1 June 2002

journal article
Published by MIT Press in Neural Computation

Vol. 14 (6), 1347-1369
https://doi.org/10.1162/089976602753712972

Abstract

We propose a modular reinforcement learning architecture for nonlinear, nonstationary control tasks, which we call multiple model-based reinforcement learning (MMRL). The basic idea is to decompose a complex task into multiple domains in space and time based on the predictability of the environmental dynamics. The system is composed of multiple modules, each of which consists of a state prediction model and a reinforcement learning controller. The "responsibility signal," which is given by the softmax function of the prediction errors, is used to weight the outputs of multiple modules, as well as to gate the learning of the prediction models and the reinforcement learning controllers. We formulate MMRL for both discrete-time, finite-state case and continuous-time, continuous-state case. The performance of MMRL was demonstrated for discrete case in a nonstationary hunting task in a grid world and for continuous case in a nonlinear, nonstationary control task of swinging up a pendulum with variable physical parameters.

Keywords

This publication has 17 references indexed in Scilit:

MOSAIC Model for Sensorimotor Learning and Control
Neural Computation, 2001
Acquisition of stand-up behavior by a real robot using hierarchical reinforcement learning
Robotics and Autonomous Systems, 2001
Reinforcement Learning in Continuous Time and Space
Neural Computation, 2000
Human cerebellar activity reflecting an acquired internal model of a new tool
Nature, 2000
Annealed Competition of Experts for a Segmentation and Classification of Switching Dynamics
Neural Computation, 1996
Adaptation and learning using multiple models, switching, and tuning
IEEE Control Systems, 1995
Recognition of manipulated objects by motor learning with modular architecture networks
Neural Networks, 1993
Transfer of learning by composing solutions of elemental sequential tasks
Machine Learning, 1992
Adaptive Mixtures of Local Experts
Neural Computation, 1991
Neuronlike adaptive elements that can solve difficult learning control problems
IEEE Transactions on Systems, Man, and Cybernetics, 1983

Cited by 328 articles