A Hybrid Reinforcement Learning Approach for Cargo Delivery by Autonomous Drone

dc.contributor.authorKarakose, Ebru
dc.contributor.authorBayraktar, Batuhan
dc.date.accessioned2026-08-12T15:30:42Z
dc.date.issued2025
dc.departmentFırat Üniversitesi
dc.description.abstractThe use of drones, particularly in the transportation sector and cargo delivery, is among the challenging and limited issues that attract significant attention and focus. In this study, a drone operates in a simulation environment created with Unreal Engine software, operating from the center of the map without any external information, not even route information, and delivers cargo completely autonomously. The drone’s missions include overcoming obstacles, remaining unaffected by weather conditions, finding the cargo vehicle, and delivering the cargo to its intended recipient. Three different algorithms, together with RGB and depth cameras, were used for cargo transportation and navigation purposes in an autonomously moving drone. Six different combinations were created, and comparisons were made across a variety of variables. Each combination was trained for 150,000 steps and evaluated against predetermined metrics. The drone was trained using reinforcement learning algorithms such as DQN, PPO, and hybrid Joint-DQN algorithms, and the LSTM algorithm was also used for memory. These algorithms were tested and compared in the simulation environment. Additionally, RGB and depth cameras were integrated into the drone, and each algorithm was run and evaluated separately using the RGB and depth cameras. In the system, the drone earns positive points as it moves toward the target and receives negative points when it moves in the opposite direction. If the drone crashes into an obstacle, the simulation restarts. The results showed that the algorithms first learned to overcome obstacles and then found the correct path. Given sufficient learning time, the drone successfully completed its mission. Furthermore, when the models were evaluated in terms of performance, the DQN-RGB model was identified as the fastest learning model, with the PPO algorithms lagging behind all other models. As a result, it was noted that although the proposed “Joint” layer slows down the learning rate, it produces a more stable and efficient model in the long run.
dc.identifier.doi10.62520/fujece.1652790
dc.identifier.endpage603
dc.identifier.issn2822-2881
dc.identifier.issue3
dc.identifier.startpage580
dc.identifier.trdizinid1351795
dc.identifier.urihttps://doi.org/10.62520/fujece.1652790
dc.identifier.urihttps://search.trdizin.gov.tr/tr/yayin/detay/1351795
dc.identifier.urihttps://hdl.handle.net/11508/32991
dc.identifier.volume4
dc.indekslendigikaynakTR-Dizin
dc.language.isoen
dc.relation.ispartofFirat University journal of experimental and computational engineering (Online)
dc.relation.publicationcategoryMakale - Ulusal Hakemli Dergi - Kurum Öğretim Elemanı
dc.relation.tubitakinfo:eu-repo/grantAgreement/TUBITAK//
dc.rightsinfo:eu-repo/semantics/openAccess
dc.snmzKA_TR-Dizin_20260511
dc.subjectReinforcement learning
dc.subjectAutonomous drone
dc.subjectDepth camera
dc.subjectCargo delivery
dc.subjectRGB camera
dc.titleA Hybrid Reinforcement Learning Approach for Cargo Delivery by Autonomous Drone
dc.typeArticle

Dosyalar