Enhancing UAV Aerial Docking: A Hybrid Approach Combining Offline and Online Reinforcement Learning

被引:0
|
作者
Feng, Yuting [1 ]
Yang, Tao [1 ]
Yu, Yushu [1 ]
机构
[1] Beijing Inst Technol, Sch Mechatron Engn, Beijing 100081, Peoples R China
基金
中国国家自然科学基金;
关键词
uav aerial docking; offline reinforcement learning; online reinforcement learning; INTERNAL DYNAMICS; QUADROTORS; PLATFORM;
D O I
10.3390/drones8050168
中图分类号
TP7 [遥感技术];
学科分类号
081102 ; 0816 ; 081602 ; 083002 ; 1404 ;
摘要
In our study, we explore the task of performing docking maneuvers between two unmanned aerial vehicles (UAVs) using a combination of offline and online reinforcement learning (RL) methods. This task requires a UAV to accomplish external docking while maintaining stable flight control, representing two distinct types of objectives at the task execution level. Direct online RL training could lead to catastrophic forgetting, resulting in training failure. To overcome these challenges, we design a rule-based expert controller and accumulate an extensive dataset. Based on this, we concurrently design a series of rewards and train a guiding policy through offline RL. Then, we conduct comparative verification on different RL methods, ultimately selecting online RL to fine-tune the model trained offline. This strategy effectively combines the efficiency of offline RL with the exploratory capabilities of online RL. Our approach improves the success rate of the UAV's aerial docking task, increasing it from 40% under the expert policy to 95%.
引用
收藏
页数:18
相关论文
共 50 条
  • [1] Learning Aerial Docking via Offline-to-Online Reinforcement Learning
    Tao, Yang
    Feng Yuting
    Yu, Yushu
    2024 4TH INTERNATIONAL CONFERENCE ON COMPUTER, CONTROL AND ROBOTICS, ICCCR 2024, 2024, : 305 - 309
  • [2] Hybrid Online and Offline Reinforcement Learning for Tibetan Jiu Chess
    Li, Xiali
    Lv, Zhengyu
    Wu, Licheng
    Zhao, Yue
    Xu, Xiaona
    COMPLEXITY, 2020, 2020
  • [3] Hybrid Offline/Online Optimization for Energy Management via Reinforcement Learning
    Silvestri, Mattia
    De Filippo, Allegra
    Ruggeri, Federico
    Lombardi, Michele
    INTEGRATION OF CONSTRAINT PROGRAMMING, ARTIFICIAL INTELLIGENCE, AND OPERATIONS RESEARCH, CPAIOR 2022, 2022, 13292 : 358 - 373
  • [4] Control of Hybrid Electric Vehicle Powertrain Using Offline-Online Hybrid Reinforcement Learning
    Yao, Zhengyu
    Yoon, Hwan-Sik
    Hong, Yang-Ki
    ENERGIES, 2023, 16 (02)
  • [5] A hybrid deep reinforcement learning approach for a proactive transshipment of fresh food in the online-offline channel system
    Lee, Junhyeok
    Shin, Youngchul
    Moon, Ilkyeong
    TRANSPORTATION RESEARCH PART E-LOGISTICS AND TRANSPORTATION REVIEW, 2024, 187
  • [6] Offline Evaluation of Online Reinforcement Learning Algorithms
    Mandel, Travis
    Liu, Yun-En
    Brunskill, Emma
    Popovic, Zoran
    THIRTIETH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, 2016, : 1926 - 1933
  • [7] Efficient Online Reinforcement Learning with Offline Data
    Ball, Philip J.
    Smith, Laura
    Kostrikov, Ilya
    Levine, Sergey
    INTERNATIONAL CONFERENCE ON MACHINE LEARNING, VOL 202, 2023, 202
  • [8] Optimum Aerial Base Station Deployment for UAV Networks: A Reinforcement Learning Approach
    Hou, Meng-Chun
    Deng, Der-Jiunn
    Wu, Chia-Ling
    2019 IEEE GLOBECOM WORKSHOPS (GC WKSHPS), 2019,
  • [9] A Minimalist Approach to Offline Reinforcement Learning
    Fujimoto, Scott
    Gu, Shixiang Shane
    ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS 34 (NEURIPS 2021), 2021, 34
  • [10] Hybrid offline-online reinforcement learning for obstacle avoidance in autonomous underwater vehicles
    Zhao, Jintao
    Liu, Tao
    Huang, Junhao
    SHIPS AND OFFSHORE STRUCTURES, 2024,