DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations

Top-performing Model-Based Reinforcement Learning (MBRL) agents, such as Dreamer, learn the world model by reconstructing the image observations. Hence, they often fail to discard task-irrelevant details and struggle to handle visual distractions. To address this issue, previous work has proposed to...

Ausführliche Beschreibung

Gespeichert in:
Bibliographische Detailangaben
Veröffentlicht in:arXiv.org 2021-10
Hauptverfasser: Deng, Fei, Jang, Ingook, Ahn, Sungjin
Format: Artikel
Sprache:eng
Schlagworte:
Online-Zugang:Volltext
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
container_end_page
container_issue
container_start_page
container_title arXiv.org
container_volume
creator Deng, Fei
Jang, Ingook
Ahn, Sungjin
description Top-performing Model-Based Reinforcement Learning (MBRL) agents, such as Dreamer, learn the world model by reconstructing the image observations. Hence, they often fail to discard task-irrelevant details and struggle to handle visual distractions. To address this issue, previous work has proposed to contrastively learn the world model, but the performance tends to be inferior in the absence of distractions. In this paper, we seek to enhance robustness to distractions for MBRL agents. Specifically, we consider incorporating prototypical representations, which have yielded more accurate and robust results than contrastive approaches in computer vision. However, it remains elusive how prototypical representations can benefit temporal dynamics learning in MBRL, since they treat each image independently without capturing temporal structures. To this end, we propose to learn the prototypes from the recurrent states of the world model, thereby distilling temporal structures from past observations and actions into the prototypes. The resulting model, DreamerPro, successfully combines Dreamer with prototypes, making large performance gains on the DeepMind Control suite both in the standard setting and when there are complex background distractions. Code available at https://github.com/fdeng18/dreamer-pro .
format Article
fullrecord <record><control><sourceid>proquest</sourceid><recordid>TN_cdi_proquest_journals_2587505062</recordid><sourceformat>XML</sourceformat><sourcesystem>PC</sourcesystem><sourcerecordid>2587505062</sourcerecordid><originalsourceid>FETCH-proquest_journals_25875050623</originalsourceid><addsrcrecordid>eNqNjEsKwjAUAIMgWLR3CLgOxNS0xaWf4kJBxH0J7aumtEl9SRFvbwQP4GoWM8yERCJJVixfCzEjsXMt51ykmZAyiUi9R1A94AXthl6hssZ5HCuvrWEFAtCzraFjW-WgDl6bxmIFPRhPT6DQaHOnL-0fNAy89e9BV6oL4YDgQqS-I7cg00Z1DuIf52RZHG67IxvQPkdwvmztiCaoUsg8k1zyVCT_VR-Ps0df</addsrcrecordid><sourcetype>Aggregation Database</sourcetype><iscdi>true</iscdi><recordtype>article</recordtype><pqid>2587505062</pqid></control><display><type>article</type><title>DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations</title><source>Free E- Journals</source><creator>Deng, Fei ; Jang, Ingook ; Ahn, Sungjin</creator><creatorcontrib>Deng, Fei ; Jang, Ingook ; Ahn, Sungjin</creatorcontrib><description>Top-performing Model-Based Reinforcement Learning (MBRL) agents, such as Dreamer, learn the world model by reconstructing the image observations. Hence, they often fail to discard task-irrelevant details and struggle to handle visual distractions. To address this issue, previous work has proposed to contrastively learn the world model, but the performance tends to be inferior in the absence of distractions. In this paper, we seek to enhance robustness to distractions for MBRL agents. Specifically, we consider incorporating prototypical representations, which have yielded more accurate and robust results than contrastive approaches in computer vision. However, it remains elusive how prototypical representations can benefit temporal dynamics learning in MBRL, since they treat each image independently without capturing temporal structures. To this end, we propose to learn the prototypes from the recurrent states of the world model, thereby distilling temporal structures from past observations and actions into the prototypes. The resulting model, DreamerPro, successfully combines Dreamer with prototypes, making large performance gains on the DeepMind Control suite both in the standard setting and when there are complex background distractions. Code available at https://github.com/fdeng18/dreamer-pro .</description><identifier>EISSN: 2331-8422</identifier><language>eng</language><publisher>Ithaca: Cornell University Library, arXiv.org</publisher><subject>Computer vision ; Distillation ; Image reconstruction ; Learning ; Prototypes ; Representations</subject><ispartof>arXiv.org, 2021-10</ispartof><rights>2021. This work is published under http://arxiv.org/licenses/nonexclusive-distrib/1.0/ (the “License”). Notwithstanding the ProQuest Terms and Conditions, you may use this content in accordance with the terms of the License.</rights><oa>free_for_read</oa><woscitedreferencessubscribed>false</woscitedreferencessubscribed></display><links><openurl>$$Topenurl_article</openurl><openurlfulltext>$$Topenurlfull_article</openurlfulltext><thumbnail>$$Tsyndetics_thumb_exl</thumbnail><link.rule.ids>777,781</link.rule.ids></links><search><creatorcontrib>Deng, Fei</creatorcontrib><creatorcontrib>Jang, Ingook</creatorcontrib><creatorcontrib>Ahn, Sungjin</creatorcontrib><title>DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations</title><title>arXiv.org</title><description>Top-performing Model-Based Reinforcement Learning (MBRL) agents, such as Dreamer, learn the world model by reconstructing the image observations. Hence, they often fail to discard task-irrelevant details and struggle to handle visual distractions. To address this issue, previous work has proposed to contrastively learn the world model, but the performance tends to be inferior in the absence of distractions. In this paper, we seek to enhance robustness to distractions for MBRL agents. Specifically, we consider incorporating prototypical representations, which have yielded more accurate and robust results than contrastive approaches in computer vision. However, it remains elusive how prototypical representations can benefit temporal dynamics learning in MBRL, since they treat each image independently without capturing temporal structures. To this end, we propose to learn the prototypes from the recurrent states of the world model, thereby distilling temporal structures from past observations and actions into the prototypes. The resulting model, DreamerPro, successfully combines Dreamer with prototypes, making large performance gains on the DeepMind Control suite both in the standard setting and when there are complex background distractions. Code available at https://github.com/fdeng18/dreamer-pro .</description><subject>Computer vision</subject><subject>Distillation</subject><subject>Image reconstruction</subject><subject>Learning</subject><subject>Prototypes</subject><subject>Representations</subject><issn>2331-8422</issn><fulltext>true</fulltext><rsrctype>article</rsrctype><creationdate>2021</creationdate><recordtype>article</recordtype><sourceid>ABUWG</sourceid><sourceid>AFKRA</sourceid><sourceid>AZQEC</sourceid><sourceid>BENPR</sourceid><sourceid>CCPQU</sourceid><sourceid>DWQXO</sourceid><recordid>eNqNjEsKwjAUAIMgWLR3CLgOxNS0xaWf4kJBxH0J7aumtEl9SRFvbwQP4GoWM8yERCJJVixfCzEjsXMt51ykmZAyiUi9R1A94AXthl6hssZ5HCuvrWEFAtCzraFjW-WgDl6bxmIFPRhPT6DQaHOnL-0fNAy89e9BV6oL4YDgQqS-I7cg00Z1DuIf52RZHG67IxvQPkdwvmztiCaoUsg8k1zyVCT_VR-Ps0df</recordid><startdate>20211027</startdate><enddate>20211027</enddate><creator>Deng, Fei</creator><creator>Jang, Ingook</creator><creator>Ahn, Sungjin</creator><general>Cornell University Library, arXiv.org</general><scope>8FE</scope><scope>8FG</scope><scope>ABJCF</scope><scope>ABUWG</scope><scope>AFKRA</scope><scope>AZQEC</scope><scope>BENPR</scope><scope>BGLVJ</scope><scope>CCPQU</scope><scope>DWQXO</scope><scope>HCIFZ</scope><scope>L6V</scope><scope>M7S</scope><scope>PIMPY</scope><scope>PQEST</scope><scope>PQQKQ</scope><scope>PQUKI</scope><scope>PRINS</scope><scope>PTHSS</scope></search><sort><creationdate>20211027</creationdate><title>DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations</title><author>Deng, Fei ; Jang, Ingook ; Ahn, Sungjin</author></sort><facets><frbrtype>5</frbrtype><frbrgroupid>cdi_FETCH-proquest_journals_25875050623</frbrgroupid><rsrctype>articles</rsrctype><prefilter>articles</prefilter><language>eng</language><creationdate>2021</creationdate><topic>Computer vision</topic><topic>Distillation</topic><topic>Image reconstruction</topic><topic>Learning</topic><topic>Prototypes</topic><topic>Representations</topic><toplevel>online_resources</toplevel><creatorcontrib>Deng, Fei</creatorcontrib><creatorcontrib>Jang, Ingook</creatorcontrib><creatorcontrib>Ahn, Sungjin</creatorcontrib><collection>ProQuest SciTech Collection</collection><collection>ProQuest Technology Collection</collection><collection>Materials Science &amp; Engineering Collection</collection><collection>ProQuest Central (Alumni Edition)</collection><collection>ProQuest Central UK/Ireland</collection><collection>ProQuest Central Essentials</collection><collection>ProQuest Central</collection><collection>Technology Collection</collection><collection>ProQuest One Community College</collection><collection>ProQuest Central Korea</collection><collection>SciTech Premium Collection</collection><collection>ProQuest Engineering Collection</collection><collection>Engineering Database</collection><collection>Publicly Available Content Database</collection><collection>ProQuest One Academic Eastern Edition (DO NOT USE)</collection><collection>ProQuest One Academic</collection><collection>ProQuest One Academic UKI Edition</collection><collection>ProQuest Central China</collection><collection>Engineering Collection</collection></facets><delivery><delcategory>Remote Search Resource</delcategory><fulltext>fulltext</fulltext></delivery><addata><au>Deng, Fei</au><au>Jang, Ingook</au><au>Ahn, Sungjin</au><format>book</format><genre>document</genre><ristype>GEN</ristype><atitle>DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations</atitle><jtitle>arXiv.org</jtitle><date>2021-10-27</date><risdate>2021</risdate><eissn>2331-8422</eissn><abstract>Top-performing Model-Based Reinforcement Learning (MBRL) agents, such as Dreamer, learn the world model by reconstructing the image observations. Hence, they often fail to discard task-irrelevant details and struggle to handle visual distractions. To address this issue, previous work has proposed to contrastively learn the world model, but the performance tends to be inferior in the absence of distractions. In this paper, we seek to enhance robustness to distractions for MBRL agents. Specifically, we consider incorporating prototypical representations, which have yielded more accurate and robust results than contrastive approaches in computer vision. However, it remains elusive how prototypical representations can benefit temporal dynamics learning in MBRL, since they treat each image independently without capturing temporal structures. To this end, we propose to learn the prototypes from the recurrent states of the world model, thereby distilling temporal structures from past observations and actions into the prototypes. The resulting model, DreamerPro, successfully combines Dreamer with prototypes, making large performance gains on the DeepMind Control suite both in the standard setting and when there are complex background distractions. Code available at https://github.com/fdeng18/dreamer-pro .</abstract><cop>Ithaca</cop><pub>Cornell University Library, arXiv.org</pub><oa>free_for_read</oa></addata></record>
fulltext fulltext
identifier EISSN: 2331-8422
ispartof arXiv.org, 2021-10
issn 2331-8422
language eng
recordid cdi_proquest_journals_2587505062
source Free E- Journals
subjects Computer vision
Distillation
Image reconstruction
Learning
Prototypes
Representations
title DreamerPro: Reconstruction-Free Model-Based Reinforcement Learning with Prototypical Representations
url https://sfx.bib-bvb.de/sfx_tum?ctx_ver=Z39.88-2004&ctx_enc=info:ofi/enc:UTF-8&ctx_tim=2025-01-20T15%3A25%3A50IST&url_ver=Z39.88-2004&url_ctx_fmt=infofi/fmt:kev:mtx:ctx&rfr_id=info:sid/primo.exlibrisgroup.com:primo3-Article-proquest&rft_val_fmt=info:ofi/fmt:kev:mtx:book&rft.genre=document&rft.atitle=DreamerPro:%20Reconstruction-Free%20Model-Based%20Reinforcement%20Learning%20with%20Prototypical%20Representations&rft.jtitle=arXiv.org&rft.au=Deng,%20Fei&rft.date=2021-10-27&rft.eissn=2331-8422&rft_id=info:doi/&rft_dat=%3Cproquest%3E2587505062%3C/proquest%3E%3Curl%3E%3C/url%3E&disable_directlink=true&sfx.directlink=off&sfx.report_link=0&rft_id=info:oai/&rft_pqid=2587505062&rft_id=info:pmid/&rfr_iscdi=true