PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 10, 2026Sensors1 citationsOpen Access

Perception-Aware Cooperative Path Planning for Multi-UAV Systems in Urban Wind Fields via Deep Reinforcement Learning

View Full Paper
JDJie DingLWLinshen WangSJShuxin Jin

Key Points

  • This research aims to improve the safety and efficiency of multi-UAV operations in complex urban environments affected by wind and obstacles.
  • Developed an enhanced Dueling Double Deep Q-Network (NPD3QN) for path planning.
  • Formulated perceived environmental data into a Markov Decision Process with N-step updates.
  • Implemented a Prioritized Experience Replay mechanism to enhance training stability.
  • Achieved a reduction in total path length by approximately 11.7% compared to the standard D3QN baseline in wind-disturbed scenarios.
  • Demonstrated effective mapping of high-dimensional state perceptions to robust control commands through extensive testing.
  • Established a robust methodological foundation for autonomous multi-UAV operations in challenging environments.

Abstract

The safe deployment of multiple Unmanned Aerial Vehicles (UAVs) in complex urban environments relies heavily on accurate environmental perception and efficient cooperative path planning. However, executing multi-UAV operations in low-altitude airspaces faces severe challenges due to the dual constraints of complex building clusters and steady-state wind field disturbances. These dynamic environmental factors frequently distort sensory expectations, inducing trajectory drift and degrading policy robustness. To address these limitations, this paper proposes an enhanced Dueling Double Deep Q-Network (D3QN) algorithm, termed NPD3QN, tailored for perception-aware multi-UAV cooperative path planning. By formulating the perceived environmental data (e.g., wind speed, obstacle distances, and inter-UAV states) into a Markov Decision Process, an N-step update strategy is integrated to enhance the characterization of long-term returns. Simultaneously, an improved Prioritized Experience Replay (PER) mechanism is developed to actively filter negative experiences and assign dynamic weights to critical state-action samples, thereby significantly elevating training stability. A 3D urban kinematic environment incorporating a steady-state simulated wind field is constructed. Extensive ablation and comparative results demonstrate that NPD3QN effectively maps high-dimensional state perceptions to robust control commands. In wind-disturbed scenarios, it generates highly streamlined cooperative trajectories, reducing the total path length by approximately 11.7% compared to the standard D3QN baseline. While currently evaluated within steady-state simulated constraints, this study establishes a robust, sensor-driven methodological foundation for autonomous multi-UAV cooperative path planning in wind-disturbed airspaces.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Ding et al. (2026) studied this question.

synapsesocial.com/papers/6a002087c8f74e3340f9b54dhttps://doi.org/10.3390/s26102960
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Three-dimensional path planning algorithm for UAV based on CL-PER-ID3QN2026
  2. 2UAV Path Planning Based on Random Obstacle Training and Linear Soft Update of DRL in Dense Urban Environment2024 · 5 citations
  3. 3Multi-Head Attention Deep Q-Network with Prioritized Experience Replay for UAV Path Planning in Dynamic Environments: A Bio-Inspired Approach2026 · 2 citations
  4. 4Multi-UAV Cooperative Path Planning Method Based on an Improved MADDPG Algorithm2026 · 2 citations
  5. 5Dynamic Scene Path Planning of UAVs Based on Deep Reinforcement Learning2024 · 56 citations