RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Raileanu, Roberta; Rocktäschel, Tim

Computer Science > Machine Learning

arXiv:2002.12292 (cs)

[Submitted on 27 Feb 2020 (v1), last revised 29 Feb 2020 (this version, v2)]

Title:RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Authors:Roberta Raileanu, Tim Rocktäschel

View PDF

Abstract:Exploration in sparse reward environments remains one of the key challenges of model-free reinforcement learning. Instead of solely relying on extrinsic rewards provided by the environment, many state-of-the-art methods use intrinsic rewards to encourage exploration. However, we show that existing methods fall short in procedurally-generated environments where an agent is unlikely to visit a state more than once. We propose a novel type of intrinsic reward which encourages the agent to take actions that lead to significant changes in its learned state representation. We evaluate our method on multiple challenging procedurally-generated tasks in MiniGrid, as well as on tasks with high-dimensional observations used in prior work. Our experiments demonstrate that this approach is more sample efficient than existing exploration methods, particularly for procedurally-generated MiniGrid environments. Furthermore, we analyze the learned behavior as well as the intrinsic reward received by our agent. In contrast to previous approaches, our intrinsic reward does not diminish during the course of training and it rewards the agent substantially more for interacting with objects that it can control.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2002.12292 [cs.LG]
	(or arXiv:2002.12292v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2002.12292

Submission history

From: Roberta Raileanu [view email]
[v1] Thu, 27 Feb 2020 18:03:16 UTC (1,918 KB)
[v2] Sat, 29 Feb 2020 16:12:58 UTC (1,918 KB)

Computer Science > Machine Learning

Title:RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators