Next Article in Journal
Underwater Acoustic MAC Protocol for Multi-Objective Optimization Based on Multi-Agent Reinforcement Learning
Previous Article in Journal
Joint Optimization of Data Collection for Multi-UAV-and-IRS-Assisted IoT in Urban Scenarios
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

A Continuous Space Path Planning Method for Unmanned Aerial Vehicle Based on Particle Swarm Optimization-Enhanced Deep Q-Network

1
School of Electronic Information Engineering, Inner Mongolia University, Hohhot 010021, China
2
Center for Applied Mathematics Inner Mongolia, Hohhot 010021, China
*
Author to whom correspondence should be addressed.
Drones 2025, 9(2), 122; https://doi.org/10.3390/drones9020122
Submission received: 13 January 2025 / Revised: 4 February 2025 / Accepted: 5 February 2025 / Published: 7 February 2025

Abstract

In the field of unmanned aerial vehicle (UAV) path planning, the conventional deep Q-network (DQN) algorithm encounters the issue of action space discretization, which results in the generation of unsmooth and inefficient planned paths. To address this issue, we introduce the particle swarm optimization (PSO) algorithm into DQN to convert the discrete action space into a continuous one. This method divides the agent’s surrounding space into discrete and continuous action spaces. The PSO algorithm performs a global search in the continuous space to obtain a continuous candidate solution, while DQN learns a policy in the discrete space to obtain a discrete candidate solution. Then, the two candidate solutions are combined using a weighted vector method to determine a direction that balances global search and policy learning. Additionally, we introduce a novel feature matrix as the state space for DQN, providing more accurate environmental and positional representations. Furthermore, we incorporate a mechanism into the base prioritized experience replay (PER) and N-step updates, which combines the current temporal difference error (TD-error) with historical priorities and includes a policy entropy penalty term, thereby enhancing DQN’s ability to learn long-term dependencies. The performance of the PSO-DQN model is further improved through an enhanced ε-greedy policy and learning rate decay strategy. Simulation results and experiments using the Flightmare simulator demonstrate that the proposed method generates smoother and more efficient paths for drones, exhibiting strong robustness in complex environments.
Keywords: UAV; particle swarm optimization; deep Q-Network; path planning UAV; particle swarm optimization; deep Q-Network; path planning

Share and Cite

MDPI and ACS Style

Han, L.; Zhang, H.; An, N. A Continuous Space Path Planning Method for Unmanned Aerial Vehicle Based on Particle Swarm Optimization-Enhanced Deep Q-Network. Drones 2025, 9, 122. https://doi.org/10.3390/drones9020122

AMA Style

Han L, Zhang H, An N. A Continuous Space Path Planning Method for Unmanned Aerial Vehicle Based on Particle Swarm Optimization-Enhanced Deep Q-Network. Drones. 2025; 9(2):122. https://doi.org/10.3390/drones9020122

Chicago/Turabian Style

Han, Le, Hui Zhang, and Nan An. 2025. "A Continuous Space Path Planning Method for Unmanned Aerial Vehicle Based on Particle Swarm Optimization-Enhanced Deep Q-Network" Drones 9, no. 2: 122. https://doi.org/10.3390/drones9020122

APA Style

Han, L., Zhang, H., & An, N. (2025). A Continuous Space Path Planning Method for Unmanned Aerial Vehicle Based on Particle Swarm Optimization-Enhanced Deep Q-Network. Drones, 9(2), 122. https://doi.org/10.3390/drones9020122

Article Metrics

Back to TopTop