GRIMGEP: Learning Progress for Robust Goal Sampling in Visual Deep Reinforcement Learning - Inria - Institut national de recherche en sciences et technologies du numérique Accéder directement au contenu
Pré-Publication, Document De Travail Année : 2021

GRIMGEP: Learning Progress for Robust Goal Sampling in Visual Deep Reinforcement Learning

Résumé

Autonomous agents using novelty based goal exploration are often efficient in environments that require exploration. However, they get attracted to various forms of distracting unlearnable regions. To solve this problem, absolute learning progress (ALP) has been used in reinforcement learning agents with predefined goal features and access to expert knowledge. This work extends those concepts to unsupervised image-based goal exploration. We present the GRIMGEP framework: it provides a learned robust goal sampling prior that can be used on top of current state-of-the-art novelty seeking goal exploration approaches, enabling them to ignore noisy distracting regions while searching for novelty in the learnable regions. It clusters the goal space and estimates ALP for each cluster. These ALP estimates can then be used to detect the distracting regions, and build a prior that enables further goal sampling mechanisms to ignore them. We construct an image based environment with distractors, on which we show that wrapping current state-of-the-art goal exploration algorithms with our framework allows them to concentrate on interesting regions of the environment and drastically improve performances. The source code is available at https://sites.google.com/view/grimgep.
Fichier principal
Vignette du fichier
GRIMGEP.pdf (960.9 Ko) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)

Dates et versions

hal-03121352 , version 1 (26-01-2021)
hal-03121352 , version 2 (22-11-2021)

Identifiants

Citer

Grgur Kovač, Adrien Laversanne-Finot, Pierre-Yves Oudeyer. GRIMGEP: Learning Progress for Robust Goal Sampling in Visual Deep Reinforcement Learning. 2021. ⟨hal-03121352v1⟩

Collections

ENSTA
97 Consultations
157 Téléchargements

Altmetric

Partager

Gmail Facebook X LinkedIn More