LG AI NEApr 4, 2019

Self-Adapting Goals Allow Transfer of Predictive Models to New Tasks

arXiv:1904.02435v21.0

Originality Incremental advance

AI Analysis

This work addresses the problem of model transferability in reinforcement learning for agents, though it appears incremental as it builds on existing predictive model architectures.

The paper tackles the challenge of transferring predictive models in reinforcement learning to new tasks by extending a deep learning architecture with an evolving neural network that suggests adaptive goals, demonstrating successful transfer to different scenarios and on-line strategy adjustments.

A long-standing challenge in Reinforcement Learning is enabling agents to learn a model of their environment which can be transferred to solve other problems in a world with the same underlying rules. One reason this is difficult is the challenge of learning accurate models of an environment. If such a model is inaccurate, the agent's plans and actions will likely be sub-optimal, and likely lead to the wrong outcomes. Recent progress in model-based reinforcement learning has improved the ability for agents to learn and use predictive models. In this paper, we extend a recent deep learning architecture which learns a predictive model of the environment that aims to predict only the value of a few key measurements, which are be indicative of an agent's performance. Predicting only a few measurements rather than the entire future state of an environment makes it more feasible to learn a valuable predictive model. We extend this predictive model with a small, evolving neural network that suggests the best goals to pursue in the current state. We demonstrate that this allows the predictive model to transfer to new scenarios where goals are different, and that the adaptive goals can even adjust agent behavior on-line, changing its strategy to fit the current context.

View on arXiv PDF

Similar