Computer Science > Neural and Evolutionary Computing
[Submitted on 12 Jan 2022 (v1), last revised 16 Sep 2022 (this version, v4)]
Title:Evolutionary Action Selection for Gradient-based Policy Learning
View PDFAbstract:Evolutionary Algorithms (EAs) and Deep Reinforcement Learning (DRL) have recently been integrated to take the advantage of the both methods for better exploration and this http URL evolutionary part in these hybrid methods maintains a population of policy this http URL, existing methods focus on optimizing the parameters of policy network, which is usually high-dimensional and tricky for this http URL this paper, we shift the target of evolution from high-dimensional parameter space to low-dimensional action this http URL propose Evolutionary Action Selection-Twin Delayed Deep Deterministic Policy Gradient (EAS-TD3), a novel hybrid method of EA and this http URL EAS, we focus on optimizing the action chosen by the policy network and attempt to obtain high-quality actions to promote policy learning through an evolutionary algorithm. We conduct several experiments on challenging continuous control this http URL result shows that EAS-TD3 shows superior performance over other state-of-art methods.
Submission history
From: Yan Ma [view email][v1] Wed, 12 Jan 2022 03:31:21 UTC (11,079 KB)
[v2] Thu, 20 Jan 2022 03:20:25 UTC (11,074 KB)
[v3] Thu, 17 Feb 2022 03:27:58 UTC (11,006 KB)
[v4] Fri, 16 Sep 2022 12:32:55 UTC (11,024 KB)
References & Citations
Bibliographic and Citation Tools
Bibliographic Explorer (What is the Explorer?)
Litmaps (What is Litmaps?)
scite Smart Citations (What are Smart Citations?)
Code, Data and Media Associated with this Article
CatalyzeX Code Finder for Papers (What is CatalyzeX?)
DagsHub (What is DagsHub?)
Gotit.pub (What is GotitPub?)
Papers with Code (What is Papers with Code?)
ScienceCast (What is ScienceCast?)
Demos
Recommenders and Search Tools
Influence Flower (What are Influence Flowers?)
Connected Papers (What is Connected Papers?)
CORE Recommender (What is CORE?)
arXivLabs: experimental projects with community collaborators
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.
Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.
Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs.