Skip to main navigation Skip to search Skip to main content

A Reinforcement Learning Framework for UAV Path Planning Automation in Complex Fire and Rescue Environments

  • Yanming Chen
  • , Saeid Pourroostaei Ardakani*
  • *Corresponding author for this work

Research output: Journal PublicationArticlepeer-review

Abstract

Fire and rescue UAVs employ AI methods to automate navigation over fire zones to detect and extinguish fire flames and/or rescue fire-trapped lives through low-cost and collision-free flight routes. This paper introduces a reinforcement learning framework that allows UAVs to detect and prioritise fire-trapped targets based on their risk levels. It establishes three simultaneous objectives, including maximised power conservation, rescue success rate, and flight safety, to plan fire and rescue operations in complex environments where field size, the number of trapped targets, and fire source count vary. An extensive empirical evaluation is conducted to test and evaluate the performance of the proposal against three well-known benchmarks, including Double Deep Q-network, Advantage Actor-Critic, and Genetic Algorithm. The results demonstrate that the proposed solution outperforms the benchmarks in most circumstances, especially when the fire and rescue environment is large and complex.

Original languageEnglish
Pages (from-to)266-277
Number of pages12
JournalIEEE Transactions on Sustainable Computing
Volume11
Issue number3
DOIs
Publication statusPublished - 1 May 2026
Externally publishedYes

UN SDGs

This output contributes to the following UN Sustainable Development Goals (SDGs)

  1. SDG 7 - Affordable and Clean Energy
    SDG 7 Affordable and Clean Energy

Free Keywords

  • Q-learning
  • Reinforcement learning
  • UAV path planning
  • deep learning
  • fire and rescue

ASJC Scopus subject areas

  • Software
  • Renewable Energy, Sustainability and the Environment
  • Hardware and Architecture
  • Control and Optimization
  • Computational Theory and Mathematics

Fingerprint

Dive into the research topics of 'A Reinforcement Learning Framework for UAV Path Planning Automation in Complex Fire and Rescue Environments'. Together they form a unique fingerprint.

Cite this