DSpace at KOASAS: Visual Pretraining via Contrastive Predictive Model for Pixel-Based Reinforcement Learning

DSpace at KOASAS

College of Engineering(공과대학)School of Electrical Engineering(전기및전자공학부)EE-Journal Papers(저널논문)

Visual Pretraining via Contrastive Predictive Model for Pixel-Based Reinforcement Learning

Cited 2 time in

Cited 0 time in

Hit : 150
Download : 0

Export

Luu, Tung M. / Vu, Thang / Nguyen, Thanh / Yoo, Chang D.researcher

In an attempt to overcome the limitations of reward-driven representation learning in vision-based reinforcement learning (RL), an unsupervised learning framework referred to as the visual pretraining via contrastive predictive model (VPCPM) is proposed to learn the representations detached from the policy learning. Our method enables the convolutional encoder to perceive the underlying dynamics through a pair of forward and inverse models under the supervision of the contrastive loss, thus resulting in better representations. In experiments with a diverse set of vision control tasks, by initializing the encoders with VPCPM, the performance of state-of-the-art vision-based RL algorithms is significantly boosted, with 44% and 10% improvement for RAD and DrQ at 100 steps, respectively. In comparison to the prior unsupervised methods, the performance of VPCPM matches or outperforms all the baselines. We further demonstrate that the learned representations successfully generalize to the new tasks that share a similar observation and action space.

Publisher: MDPI

Issue Date: 2022-09

Language: English

Article Type: Article

Citation: SENSORS, v.22, no.17

ISSN: 1424-8220

DOI: 10.3390/s22176504

URI: http://hdl.handle.net/10203/298606

Appears in Collection: EE-Journal Papers(저널논문)

Files in This Item: There are no files associated with this item.

This item is cited by other documents in WoS

⊙ Detail Information in WoSⓡ	Click to see
⊙ Cited 2 items in WoS	Click to see citing articles in

Display Full Item Record

qr_code

트윗하기

KOASAS

Knowledge Service Development Team, KAIST 291 Daehak-ro, Yuseong-gu, Daejeon 34141, Republic of Korea. T. 82-42-350-4493 Email. koasas@kaist.ac.kr
Copyright © 2016. Korea Advanced Institute of Science and Technology. All Rights Reserved.

KOASAS

KOASAS

Browse

Visual Pretraining via Contrastive Predictive Model for Pixel-Based Reinforcement Learning

This item is cited by other documents in WoS

KOASAS

Communities & Collections