Toward Robust Long Range Policy Transfer

Tseng, Wei-Cheng; Lin, Jin-Siang; Feng, Yao-Min; Sun, Min

Computer Science > Machine Learning

arXiv:2103.02957 (cs)

[Submitted on 4 Mar 2021]

Title:Toward Robust Long Range Policy Transfer

Authors:Wei-Cheng Tseng, Jin-Siang Lin, Yao-Min Feng, Min Sun

View PDF

Abstract:Humans can master a new task within a few trials by drawing upon skills acquired through prior experience. To mimic this capability, hierarchical models combining primitive policies learned from prior tasks have been proposed. However, these methods fall short comparing to the human's range of transferability. We propose a method, which leverages the hierarchical structure to train the combination function and adapt the set of diverse primitive polices alternatively, to efficiently produce a range of complex behaviors on challenging new tasks. We also design two regularization terms to improve the diversity and utilization rate of the primitives in the pre-training phase. We demonstrate that our method outperforms other recent policy transfer methods by combining and adapting these reusable primitives in tasks with continuous action space. The experiment results further show that our approach provides a broader transferring range. The ablation study also shows the regularization terms are critical for long range policy transfer. Finally, we show that our method consistently outperforms other methods when the quality of the primitives varies.

Comments:	Accepted by AAAI 2021
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2103.02957 [cs.LG]
	(or arXiv:2103.02957v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2103.02957

Submission history

From: Wei-Cheng Tseng [view email]
[v1] Thu, 4 Mar 2021 11:17:03 UTC (4,758 KB)

Computer Science > Machine Learning

Title:Toward Robust Long Range Policy Transfer

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Toward Robust Long Range Policy Transfer

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators