CERESResearch Repository

Tactical planning interception enhancement using expert learning - twin delayed deep deterministic policy gradient

Loading...
Thumbnail Image

Date published

Free to read from

2025-10-21

Supervisor/s

Industry supervisor/s

Journal Title

Journal ISSN

Volume Title

Department

Course name

ISSN

2379-8858

Format

Citation

Lucotte N, Perrusquía A, Tsourdos A, et al., (2026) Tactical planning interception enhancement using expert learning - twin delayed deep deterministic policy gradient. IEEE Transactions on Intelligent Vehicles, Volume 11, Issue 1, January 2026, pp. 174-184

Abstract

The accurate interception of adversarial unmanned aerial vehicles (UAVs) is paramount for the protection of people and national facilities. Urban cities pose several challenges for target interception algorithms due to the presence of buildings and flying constraints that limit the manoeuvrability of UAVs for target interception. Deep Reinforcement Learning (DRL) algorithms have been deployed to solve the task effectively. However, the design of its inner elements such as the reward function and action distribution limits its generalisation to different environments. To solve this issue, this paper proposes a novel twin-delayed deep deterministic policy gradient (TD3) based expert learning algorithm that combines previous expert experiences with on-line learning to regularise and improve the policy learning effectively. This is done by following an action distribution algorithm that allows a learner agent to mix its own actions with expert ones for learning improvement and fast convergence. Extensive simulation studies are carried out under diverse urban cities configurations to show the robustness and high-accuracy of the proposed approach compared with traditional DRL baseline algorithms.

Description

Software description

Software language

Git repository

Keywords

46 Information and Computing Sciences, 4611 Machine Learning, 4002 Automotive engineering, 4007 Control engineering, mechatronics and robotics, 4603 Computer vision and multimedia computation, Expert learning, Artificial Potential Field, Twin-Delayed Deep deterministic policy gradient, interception, collision avoidance, urban city

DOI

Rights

Attribution 4.0 International

Funder/s

Grant number

Relationships

Relationships

Resources