Population-Guided Parallel Policy Search for Reinforcement Learning

Jung, Whiyoung; Park, Giseung; Sung, Youngchul

Computer Science > Machine Learning

arXiv:2001.02907 (cs)

[Submitted on 9 Jan 2020]

Title:Population-Guided Parallel Policy Search for Reinforcement Learning

Authors:Whiyoung Jung, Giseung Park, Youngchul Sung

View PDF

Abstract:In this paper, a new population-guided parallel learning scheme is proposed to enhance the performance of off-policy reinforcement learning (RL). In the proposed scheme, multiple identical learners with their own value-functions and policies share a common experience replay buffer, and search a good policy in collaboration with the guidance of the best policy information. The key point is that the information of the best policy is fused in a soft manner by constructing an augmented loss function for policy update to enlarge the overall search region by the multiple learners. The guidance by the previous best policy and the enlarged range enable faster and better policy search. Monotone improvement of the expected cumulative return by the proposed scheme is proved theoretically. Working algorithms are constructed by applying the proposed scheme to the twin delayed deep deterministic (TD3) policy gradient algorithm. Numerical results show that the constructed algorithm outperforms most of the current state-of-the-art RL algorithms, and the gain is significant in the case of sparse reward environment.

Comments:	Accepted to ICLR 2020
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Machine Learning (stat.ML)
Cite as:	arXiv:2001.02907 [cs.LG]
	(or arXiv:2001.02907v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2001.02907

Submission history

From: Whiyoung Jung [view email]
[v1] Thu, 9 Jan 2020 10:13:57 UTC (4,373 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2020-01

Change to browse by:

cs
cs.AI
stat
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

Youngchul Sung

export BibTeX citation

Computer Science > Machine Learning

Title:Population-Guided Parallel Policy Search for Reinforcement Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Population-Guided Parallel Policy Search for Reinforcement Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators