DSpace at KOASAS: DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations

DSpace at KOASAS

College of Engineering(공과대학)School of Computing(전산학부)CS-Conference Papers(학술회의논문)

DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations

Cited 2 time in

Cited 0 time in scopus

Hit : 64
Download : 0

Export

Deng, Fei / Ahn, Sungjin researcher / Jang, Ingook

Top-performing Model-Based Reinforcement Learning (MBRL) agents, such as Dreamer, learn the world model by reconstructing the image observations. Hence, they often fail to discard task-irrelevant details and struggle to handle visual distractions. To address this issue, previous work has proposed to contrastively learn the world model, but the performance tends to be inferior in the absence of distractions. In this paper, we seek to enhance robustness to distractions for MBRL agents. Specifically, we consider incorporating prototypical representations, which have yielded more accurate and robust results than contrastive approaches in computer vision. However, it remains elusive how prototypical representations can benefit temporal dynamics learning in MBRL, since they treat each image independently without capturing temporal structures. To this end, we propose to learn the prototypes from the recurrent states of the world model, thereby distilling temporal structures from past observations and actions into the prototypes. The resulting model, DreamerPro, successfully combines Dreamer with prototypes, making large performance gains on the DeepMind Control suite both in the standard setting and when there are complex background distractions.

Publisher: The International Conference on Machine Learning (ICML)

Issue Date: 2022-07-17

Language: English

Citation: Thirty-ninth International Conference on Machine Learning, ICML 2022

ISSN: 2640-3498

URI: http://hdl.handle.net/10203/299467

Appears in Collection: CS-Conference Papers(학술회의논문)

Files in This Item: There are no files associated with this item.

This item is cited by other documents in WoS

⊙ Detail Information in WoSⓡ	Click to see
⊙ Cited 2 items in WoS	Click to see citing articles in

Display Full Item Record

qr_code

트윗하기

KOASAS

Knowledge Service Development Team, KAIST 291 Daehak-ro, Yuseong-gu, Daejeon 34141, Republic of Korea. T. 82-42-350-4493 Email. koasas@kaist.ac.kr
Copyright © 2016. Korea Advanced Institute of Science and Technology. All Rights Reserved.

KOASAS

KOASAS

Browse

DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations

This item is cited by other documents in WoS

KOASAS

Communities & Collections