Adaptive evolution strategy with ensemble of mutations for Reinforcement Learning

被引：18

作者：

Ajani, Oladayo S. ^{[1
]}

Mallipeddi, Rammohan ^{[1
]}

机构：

[1] Kyungpook Natl Univ, Dept Artificial Intelligence, Daegu, South Korea

来源：

KNOWLEDGE-BASED SYSTEMS | 2022年 / 245卷

基金：

新加坡国家研究基金会;

关键词：

Evolution strategy; Reinforcement Learning; Ensemble; Mutation strategy; Black-box optimization; INITIALIZATION; ADAPTATION;

D O I：

10.1016/j.knosys.2022.108624

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

Evolving the weights of learning networks through evolutionary computation (neuroevolution) has proven scalable over a range of challenging Reinforcement Learning (RL) control tasks. However, similar to most black-box optimization problems, existing neuroevolution approaches require an additional adaptation process to effectively balance exploration and exploitation through the selection of sensitive hyper-parameters throughout the evolution process. Therefore, these methods are often plagued by the computation complexities of such adaptation processes which often rely on a number of sophisticatedly formulated strategy parameters. In this paper, Evolution Strategy (ES) with a simple yet efficient ensemble of mutation strategies is proposed. Specifically, two distinct mutation strategies coexist throughout the evolution process where each strategy is associated with its own population subset. Consequently, elites for generating a population of offspring are realized by co-evaluation of the combined population. Experiments on testbed of six (6) black-box optimization problems which are generated using a classical control problem and six (6) proven continuous RL agents demonstrate the efficiency of the proposed method in terms of faster convergence and scalability than the canonical ES. Furthermore, the proposed Adaptive Ensemble ES (AEES) shows an average of 5 - 10000x and 10 100x better sample complexity in low and high dimension problems, respectively than their associated base DRL agents.

引用

页数：8

共 51 条

[1]

[Anonymous], STUDIES COMPUTATIONA, DOI DOI 10.1007/978-3-319-05029-4_7

[2]

[Anonymous], IEEE Transactions on Intelligent Transportation Systems

[3]

[Anonymous], 2021, SPRINGER OPTIMIZATIO

[4]

[Anonymous], 1995, Evolution and optimum seeking

[5]

Beyer H.-G., 2004, IEEE T EVOLUT COMPUT, V1, P3

[6] Simplify Your Covariance Matrix Adaptation Evolution Strategy [J].

Beyer, Hans-Georg ;

Sendhoff, Bernhard .

IEEE TRANSACTIONS ON EVOLUTIONARY COMPUTATION, 2017, 21 (05) :746-759

[7]

Brockman Greg, 2016, arXiv

[8]

Chen ZF, 2019, PROCEEDINGS OF THE TWENTY-EIGHTH INTERNATIONAL JOINT CONFERENCE ON ARTIFICIAL INTELLIGENCE, P2130

[9]

Chrabaszcz P., 2018, PREPRINT

[10]

Conti Edoardo, 2018, P 32 INT C NEUR INF, P5032

← 1 2 3 4 5 6 →