DSpace at KOASAS: Utilizing Skipped Frames in Action Repeats for Improving Sample Efficiency in Reinforcement Learning

DSpace at KOASAS

College of Engineering(공과대학)School of Electrical Engineering(전기및전자공학부)EE-Journal Papers(저널논문)

Utilizing Skipped Frames in Action Repeats for Improving Sample Efficiency in Reinforcement Learning

Cited 1 time in

Cited 0 time in

Hit : 152
Download : 0

Export

DC Field	Value	Language
dc.contributor.author	Luu, Tung M.	ko
dc.contributor.author	Nguyen, Thanh	ko
dc.contributor.author	Vu, Thang	ko
dc.contributor.author	Yoo, Chang-Dong	ko
dc.date.accessioned	2022-06-26T01:02:00Z	-
dc.date.available	2022-06-26T01:02:00Z	-
dc.date.created	2022-06-25	-
dc.date.created	2022-06-25	-
dc.date.created	2022-06-25	-
dc.date.created	2022-06-25	-
dc.date.created	2022-06-25	-
dc.date.issued	2022	-
dc.identifier.citation	IEEE ACCESS, v.10, pp.64965 - 64975	-
dc.identifier.issn	2169-3536	-
dc.identifier.uri	http://hdl.handle.net/10203/297080	-
dc.description.abstract	Action repeat has become the de-facto mechanism in deep reinforcement learning (RL) for stabilizing training and enhancing exploration. Here, the action is taken at the action-decision point and is executed repeatedly for a designated number of times until the next decision point. Although showing several advantages, in this mechanism, the intermediate states which stem from repeated actions are discarded in training agents, causing sample inefficiency. To utilize the discarded states as training data is nontrivial as the action, which causes the transition between these states, is unavailable. This paper proposes to infer the action at the intermediate states via an inverse dynamic model. The proposed method is simple and easily incorporated into the existing off-policy RL algorithms - integrating the proposed method with SAC shows consistent improvement across various tasks.	-
dc.language	English	-
dc.publisher	IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC	-
dc.title	Utilizing Skipped Frames in Action Repeats for Improving Sample Efficiency in Reinforcement Learning	-
dc.type	Article	-
dc.identifier.wosid	000815504400001	-
dc.identifier.scopusid	2-s2.0-85132792201	-
dc.type.rims	ART	-
dc.citation.volume	10	-
dc.citation.beginningpage	64965	-
dc.citation.endingpage	64975	-
dc.citation.publicationname	IEEE ACCESS	-
dc.identifier.doi	10.1109/access.2022.3182107	-
dc.contributor.localauthor	Yoo, Chang-Dong	-
dc.contributor.nonIdAuthor	Luu, Tung M.	-
dc.contributor.nonIdAuthor	Nguyen, Thanh	-
dc.description.isOpenAccess	N	-
dc.type.journalArticle	Article	-
dc.subject.keywordAuthor	Task analysis	-
dc.subject.keywordAuthor	Training	-
dc.subject.keywordAuthor	Heuristic algorithms	-
dc.subject.keywordAuthor	Benchmark testing	-
dc.subject.keywordAuthor	Data models	-
dc.subject.keywordAuthor	Training data	-
dc.subject.keywordAuthor	Robots	-
dc.subject.keywordAuthor	Action repeat mechanism	-
dc.subject.keywordAuthor	off-policy reinforcement learning	-
dc.subject.keywordAuthor	reinforcement learning	-
dc.subject.keywordAuthor	sample efficiency	-
dc.subject.keywordPlus	LEVEL	-

Appears in Collection: EE-Journal Papers(저널논문)

Files in This Item: There are no files associated with this item.

This item is cited by other documents in WoS

⊙ Detail Information in WoSⓡ	Click to see
⊙ Cited 1 items in WoS	Click to see citing articles in

Display Simple Item Record

qr_code

트윗하기

KOASAS

Knowledge Service Development Team, KAIST 291 Daehak-ro, Yuseong-gu, Daejeon 34141, Republic of Korea. T. 82-42-350-4493 Email. koasas@kaist.ac.kr
Copyright © 2016. Korea Advanced Institute of Science and Technology. All Rights Reserved.

KOASAS

KOASAS

Browse

Utilizing Skipped Frames in Action Repeats for Improving Sample Efficiency in Reinforcement Learning

This item is cited by other documents in WoS

KOASAS

Communities & Collections