Skip to main navigation Skip to search Skip to main content

Swarm unmanned surface vehicle encirclement task with multi-agent reinforcement learning

Research output: Contribution to journalConference articlepeer-review

Abstract

Swarm unmanned surface vehicles (USVs) have become a key technology for future maritime defense due to their significant capability for cooperative and autonomous operations. Based on swarm coordination strategies, encirclement tasks make a significant contribution to target containment, interception, and area protection. This study investigates a multi-agent reinforcement learning (MARL) approach for the swarm USV encirclement task, comparing two algorithms: the basic Multi-Agent Proximal Policy Optimization (MAPPO) and its recurrent extension, MAPPO-LSTM. These algorithms were trained in Unity 3D simulation platform using ML-Agents toolkit. Three defenders cooperatively encircle a target in an adapted real-map water environment. The performance evaluation is conducted using four metrics (cumulative reward, angular coverage, maximum angular gap, and rotation number of encirclement). Experimental results show that MAPPO-LSTM reaches better cumulative rewards and improved temporal stability, while the basic MAPPO model produces broader spatial coverage and tighter angular formation. The use of LSTM improves motion smoothness and coordination through temporal memory, resulting in more consistent encirclement behavior. These findings highlight the trade-off between spatial completeness and temporal coherence in USV swarm encirclement and emphasize the potential of the MARL framework for smart maritime defense applications.

Original languageEnglish
Pages (from-to)468-475
Number of pages8
JournalTransportation Research Procedia
Volume97
DOIs
StatePublished - 2026
Event13th International Conference on Transport Survey Methods, 2026 - Danang, Viet Nam
Duration: 30 Mar 20254 Apr 2025

Bibliographical note

Publisher Copyright:
Copyright © 2026. Published by Elsevier B.V.

Keywords

  • Encirclement
  • LSTM
  • MAPPO
  • Maritime Defense
  • Multi-Agent Reinforcement Learning
  • Swarm USV

ASJC Scopus subject areas

  • Transportation

Fingerprint

Dive into the research topics of 'Swarm unmanned surface vehicle encirclement task with multi-agent reinforcement learning'. Together they form a unique fingerprint.

Cite this