CERESResearch Repository

Temporal-augmented observation for navigation of unmanned aerial vehicles: a recurrent reinforcement learning architecture

Loading...
Thumbnail Image

Date published

Free to read from

2026-07-29

Supervisor/s

Industry supervisor/s

Journal Title

Journal ISSN

Volume Title

Department

Course name

ISSN

Format

Citation

Gemignani G, Perrusquía A, Tsourdos A, Pollini L. (2026) Temporal-augmented observation for navigation of unmanned aerial vehicles: a recurrent reinforcement learning architecture. In: Proceeding of the 2026 International Conference on Unmanned Aircraft Systems (ICUAS), 15-18 Jun 2026, Corfu, Greece, pp. 496-502

Abstract

In aerial robotics, data-driven Reinforcement Learning (RL) approaches have proven highly effective for obstacle avoidance and goal-directed navigation, especially when operating on high-dimensional sensor data that provide only partial, local information about the environment. Such limited observability, combined with irregularly shaped obstacles, poses significant challenges for reactive control policies that rely solely on instantaneous observations. To address these issues, this paper introduces a Twin Deep Deterministic Policy Gradient (TD3)-based algorithm that leverages explicit Temporal Augmentation of the Observation space (TAO-TD3). The proposed method preserves the simplicity of the original TD3 framework by augmenting the observation with a short history of past states and incorporating a lightweight recurrent network, without requiring changes to the TD3 training paradigm. Extensive simulations across diverse environmental topographies and irregular obstacle shapes demonstrate that the proposed approach nearly halves the collision rate and improves overall navigation success compared to feedforward RL-based architectures.

Description

Software description

Software language

Git repository

Keywords

46 Information and Computing Sciences, 4602 Artificial Intelligence, 4611 Machine Learning, Machine Learning and Artificial Intelligence

DOI

Rights

Attribution 4.0 International

Funder/s

Project co-funded by the European Union – Next Generation Eu - under the National Recovery and Resilience Plan (NRRP), Mission 4 Component 1 Investment 4.1 -Decree No. 118 (2nd March 2023) of Italian Ministry of University and Research - Concession Decree No. 2333 (22nd December 2023) of the Italian Ministry of University and Research, Project code D93C23000450005, within the Italian National Program PhD Programme in Autonomous Systems (DAuSy).

Grant number

Relationships

Relationships

Resources