In this paper, we study a joint detection, mapping and navigation problem for a single unmanned aerial vehicle (UAV) equipped with a low complexity radar and flying in an unknown environment. The goal is to optimize its trajectory with the purpose of maximizing the mapping accuracy and, at the same time, to avoid areas where measurements might not be sufficiently informative from the perspective of a target detection. This problem is formulated as a Markov decision process (MDP) where the UAV is an agent that runs either a state estimator for target detection and for environment mapping, and a reinforcement learning (RL) algorithm to infer its own policy of navigation (i.e., the control law). Numerical results show the feasibility of the proposed idea, highlighting the UAV's capability of autonomously exploring areas with high probability of target detection while reconstructing the surrounding environment.
Guerra A., Guidi F., Dardari D., Djuric P.M. (2020). Reinforcement Learning for UAV Autonomous Navigation, Mapping and Target Detection. Piscataway, N.J. : IEEE-Institute of Electrical and Electronics Engineers [10.1109/PLANS46316.2020.9110163].
Reinforcement Learning for UAV Autonomous Navigation, Mapping and Target Detection
Guerra A.;Guidi F.;Dardari D.;
2020
Abstract
In this paper, we study a joint detection, mapping and navigation problem for a single unmanned aerial vehicle (UAV) equipped with a low complexity radar and flying in an unknown environment. The goal is to optimize its trajectory with the purpose of maximizing the mapping accuracy and, at the same time, to avoid areas where measurements might not be sufficiently informative from the perspective of a target detection. This problem is formulated as a Markov decision process (MDP) where the UAV is an agent that runs either a state estimator for target detection and for environment mapping, and a reinforcement learning (RL) algorithm to infer its own policy of navigation (i.e., the control law). Numerical results show the feasibility of the proposed idea, highlighting the UAV's capability of autonomously exploring areas with high probability of target detection while reconstructing the surrounding environment.File | Dimensione | Formato | |
---|---|---|---|
IONGuerra.pdf
accesso riservato
Tipo:
Versione (PDF) editoriale
Licenza:
Licenza per accesso riservato
Dimensione
664.98 kB
Formato
Adobe PDF
|
664.98 kB | Adobe PDF | Visualizza/Apri Contatta l'autore |
postprint_IEEE_ION.pdf
Open Access dal 09/12/2020
Tipo:
Postprint
Licenza:
Licenza per accesso libero gratuito
Dimensione
941.48 kB
Formato
Adobe PDF
|
941.48 kB | Adobe PDF | Visualizza/Apri |
I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.