PaperScope
LIVE · 2026-09-10 05:40 UTC

Multi-Agent Reinforcement Learning for Autonomous UAV Exploration in Wildfire Response

Caden Chandra, Jerry Ng

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.10433 v1
Category
Submitted
2026-09-09

Abstract

This study develops a deep reinforcement learning framework for training Unmanned Aerial Vehicle (UAV) agents to navigate and monitor simulated wildfire environments. Results show that agents learn increasingly stable and effective behaviors over time, as demonstrated by converging loss trends, improved reward signals, and more consistent navigation patterns such as fire-boundary tracking. Overall, these findings highlight the potential of deep reinforcement learning (DRL) based UAV systems for autonomous wildfire monitoring and suggest that environmental structure and reward design influence policy effectiveness.

arXiv abs page · PDF · same-day batch