Yayınlanmış 1 Ocak 2023 | Sürüm v1
Konferans bildirisi Açık

Maintaining Connectivity for Multi-UAV Multi-Target Search Using Reinforcement Learning

  • 1. Ozyegin Univ, Elect & Elect Engn Dept, Istanbul, Turkiye

Açıklama

We propose a dynamic path planner that uses a multi-agent reinforcement learning (MARL) model with novel reward functions for multi-drone search and rescue (SAR) missions. We design a mission environment where a multi-drone team covers an area to detect randomly distributed targets and inform the ground base station (BS) by continuously forming relay chains between the targets and the BS. The training procedure of the agents includes a convolutional neural network (CNN) that uses images which represent trajectory histories and connectivity states of each environment entity such as drones, targets, BS. Agents take actions and get feedback from the environment until the mission is completed. The model is trained with multiple missions with randomized target locations. Our results show that the trained model successfully produces mission plans such that the multi-drone system searches the area efficiently while dynamically forming relay chains. The proposed dynamic method leads up to 45% better total detection and mission times in comparison to a pre-planned optimized path planner.

Dosyalar

bib-80549b0a-5e03-4e38-ac6b-9a7e86ca7aa9.txt

Dosyalar (248 Bytes)

Ad Boyut Hepisini indir
md5:8c06f4b3e6896f27e226fcc8f1c150d6
248 Bytes Ön İzleme İndir