A Reinforcement Learning Approach To Coordinate Exploration With Limited Communication In Continuous Action Games

Abdel Rodriguez,Peter Vrancx,Ricardo Grau,Ann Nowe

Knowledge Engineering Review（2016）

引用 1|浏览35

暂无评分

摘要

Learning automata are reinforcement learners belonging to the class of policy iterators. They have already been shown to exhibit nice convergence properties in a wide range of discrete action game settings. Recently, a new formulation for a continuous action reinforcement learning automata (CARLA) was proposed. In this paper, we study the behavior of these CARLA in continuous action games and propose a novel method for coordinated exploration of the joint-action space. Our method allows a team of independent learners, using CARLA, to find the optimal joint action in common interest settings. We first show that independent agents using CARLA will converge to a local optimum of the continuous action game. We then introduce a method for coordinated exploration which allows the team of agents to find the global optimum of the game. We validate our approach in a number of experiments.

查看译文

关键词

continuous action games,exploration,reinforcement learning approach,limited communication

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要