Multi-UAV Path Planning and Following Based on Multi-Agent Reinforcement Learning

被引：14

作者：

Zhao, Xiaoru ^{[1
]}

Yang, Rennong ^{[1
]}

Zhong, Liangsheng ^{[2
]}

Hou, Zhiwei ^{[2
]}

机构：

[1] Air Force Engn Univ, Air Traff Control & Nav Sch, Xian 710051, Peoples R China

[2] Sun Yat Sen Univ, Sch Syst Sci & Engn, Guangzhou 510275, Peoples R China

来源：

DRONES | 2024年 / 8卷 / 01期

关键词：

path planning; path follow; deep reinforcement learning; multi-UAV; parameter share;

D O I：

10.3390/drones8010018

中图分类号：

TP7 [遥感技术];

学科分类号：

081102 ; 0816 ; 081602 ; 083002 ; 1404 ;

摘要：

Dedicated to meeting the growing demand for multi-agent collaboration in complex scenarios, this paper introduces a parameter-sharing off-policy multi-agent path planning and the following approach. Current multi-agent path planning predominantly relies on grid-based maps, whereas our proposed approach utilizes laser scan data as input, providing a closer simulation of real-world applications. In this approach, the unmanned aerial vehicle (UAV) uses the soft actor-critic (SAC) algorithm as a planner and trains its policy to converge. This policy enables end-to-end processing of laser scan data, guiding the UAV to avoid obstacles and reach the goal. At the same time, the planner incorporates paths generated by a sampling-based method as following points. The following points are continuously updated as the UAV progresses. Multi-UAV path planning tasks are facilitated, and policy convergence is accelerated through sharing experiences among agents. To address the challenge of UAVs that are initially stationary and overly cautious near the goal, a reward function is designed to encourage UAV movement. Additionally, a multi-UAV simulation environment is established to simulate real-world UAV scenarios to support training and validation of the proposed approach. The simulation results highlight the effectiveness of the presented approach in both the training process and task performance. The presented algorithm achieves an 80% success rate to guarantee that three UAVs reach the goal points.

引用

页数：18

共 50 条

[41] Graph Convolutional Multi-Agent Reinforcement Learning for UAV Coverage Control [J].

Dai, Anna ;

Li, Rongpeng ;

Zhaot, Zhifeng ;

Zhang, Honggang .

2020 12TH INTERNATIONAL CONFERENCE ON WIRELESS COMMUNICATIONS AND SIGNAL PROCESSING (WCSP), 2020, :1106-1111

[42] Hierarchical Multi-UAV Path Planning for Urban Low Altitude Environments [J].

Lei, Haoxiang ;

Yan, Yuehao ;

Liu, Jilong ;

Han, Qiang ;

Li, Zhouguan .

IEEE ACCESS, 2024, 12 :162109-162121

[43] Bio-Inspired Multi-UAV Path Planning Heuristics: A Review [J].

Aljalaud, Faten ;

Kurdi, Heba ;

Youcef-Toumi, Kamal .

MATHEMATICS, 2023, 11 (10)

[44] Efficient multi-agent path planning [J].

Arikan, O ;

Chenney, S ;

Forsyth, DA .

COMPUTER ANIMATION AND SIMULATION 2001, 2001, :151-162

[45] Path Planning for Multi-agent Systems Using Deep Q-Networks Reinforcement Learning [J].

Alispahic, Ibrahim ;

Tahirovic, Adnan .

TOWARDS AUTONOMOUS ROBOTIC SYSTEMS, TAROS 2024, PT II, 2025, 15052 :333-344

[46] Action Correction-Enhanced Multi-Agent Reinforcement Learning for Path Planning in Urban Environments [J].

Pan, Haixia ;

Han, Linfeng ;

Yan, Jiaming ;

Liu, Ruijun .

UNMANNED SYSTEMS, 2025,

[47] Group-Based Deep Reinforcement Learning in Multi-UAV Confrontation [J].

Lie, Shengang ;

Wang, Baolai ;

Xie, Tao .

NEURAL INFORMATION PROCESSING, ICONIP 2021, PT V, 2021, 1516 :617-624

[48] Multi-UAV Cooperative Autonomous Navigation Based on Multi-agent Deep Deterministic Policy Gradient [J].

Li B. ;

Yue K.-Q. ;

Gan Z.-G. ;

Gao P.-X. .

Yuhang Xuebao/Journal of Astronautics, 2021, 42 (06) :757-765

[49] Optimizing UAV-UGV coalition operations: A hybrid clustering and multi-agent reinforcement learning approach for path planning in obstructed environment [J].

Brotee, Shamyo ;

Kabir, Farhan ;

Razzaque, Md. Abdur ;

Roy, Palash ;

Mamun-Or-Rashid, Md. ;

Hassan, Md. Rafiul ;

Hassan, Mohammad Mehedi .

AD HOC NETWORKS, 2024, 160

[50] Dynamic Multi-UAV Path Planning for Multi-Target Search and Connectivity [J].

Yanmaz, Evsen ;

Balanji, Hamid Majidi ;

Guven, Islam .

IEEE TRANSACTIONS ON VEHICULAR TECHNOLOGY, 2024, 73 (07) :10516-10528

← 1 2 3 4 5 →