PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 7, 2026Drones1 citationsOpen Access

ED-SAC Reinforcement Learning-Based Adaptive Cruise Trajectory Planning Method for UAVs in Grassland Highway Inspection Scenarios

View Full Paper
SZShuhui ZhangDCDeqi ChenWZWenhui Zhang

Key Points

  • The aim is to enhance UAV trajectory tracking and control stability in grassland highway inspection scenarios using a new RL-based method.
  • Developed an ED-SAC algorithm based on an ensemble Q-network and Soft Actor-Critic.
  • Conducted simulations using the PyBullet platform for UAV inspections.
  • Compared ED-SAC against PPO, TD3, and SAC algorithms on three trajectory scenarios.
  • ED-SAC achieved the highest mission success rate of 98.7% and a tracking error of 0.27 m under disturbance-free conditions.
  • Maintained a success rate of 96.2% under continuous random disturbances.
  • Demonstrated improved training stability and anti-disturbance capability for UAVs.

Abstract

To address the issue of traffic accidents caused by livestock crossing roads on grassland highways, this paper proposes an adaptive cruise control method for unmanned aerial vehicles (UAVs) based on an ensemble Q-network and a Soft Actor-Critic (SAC) with delayed policy updates, namely the ED-SAC algorithm. Building upon the standard SAC framework, this method introduces multiple independent Critic networks to form an ensemble Q-network, and employs a random subset minimization strategy during the calculation of target Q-values to mitigate policy bias resulting from overestimated values; simultaneously, a delayed policy update mechanism decouples the optimization processes of the Actor and Critic networks, thereby enhancing training stability and control robustness. Using the PyBullet simulation platform, this paper constructs a UAV inspection scenario on grassland roads and designs three typical test tasks: infinite loop, grid scan and spiral trajectories, to conduct comparative validation of the PPO, TD3, SAC and ED-SAC algorithms. Experimental results demonstrate that, under disturbance-free conditions, ED-SAC achieves the highest mission success rate and the lowest tracking error across all three trajectory scenarios, with an average tracking error as low as 0.27 m and a mission success rate as high as 98.7%. Under continuous random external disturbances, ED-SAC still maintains high trajectory tracking accuracy and attitude control stability, with a mission success rate reaching up to 96.2%. The results demonstrate that the proposed ED-SAC algorithm can effectively enhance the trajectory tracking accuracy, training stability and anti-disturbance capability of UAVs in complex grassland road inspection scenarios, providing a reliable intelligent control method for active grassland road inspection and traffic safety early warning.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zhang et al. (2026) studied this question.

synapsesocial.com/papers/69fbe382164b5133a91a2b65https://doi.org/10.3390/drones10050347
Ask AI
Helpful
Bookmark
Share
View Full Paper