Autonomous sailing with sim-to-real reinforcement learning

Journal Article (2026)
Author(s)

Kiki J.A. Bink (Maritime Research Institute Netherlands (MARIN), Student TU Delft)

Bülent Düz (Maritime Research Institute Netherlands (MARIN))

Gabriel D. Weymouth (TU Delft - Mechanical Engineering)

Research Group
Ship Hydromechanics
DOI related publication
https://doi.org/10.1016/j.engappai.2026.115804 Final published version
More Info
expand_more
Publication Year
2026
Language
English
Research Group
Ship Hydromechanics
Journal title
Engineering Applications of Artificial Intelligence
Volume number
182
Article number
115804
Downloads counter
17
Reuse Rights

Other than for strictly personal use, it is not permitted to download, forward or distribute the text or part of it, without the consent of the author(s) and/or copyright holder(s), unless the work is under an open content license such as Creative Commons.

Abstract

Autonomous sailing offers a sustainable alternative for reducing greenhouse gas emissions in maritime transport, aligning with global environmental targets. This study explores the application of reinforcement learning (RL) to autonomous sailing, addressing challenges in handling dynamic and unpredictable environmental conditions. Leveraging a sim-to-real transfer methodology, RL agents were trained in a simulation environment with the domain randomization technique to enhance adaptability and robustness, and tested in real-world scenarios using a robotic sailboat in the Offshore Basin at MARIN. The study quantified the reality gap between simulation and real-world environments, identifying key discrepancies in actuator latency and simulation modeling accuracy. In real-world basin experiments, the best-performing RL agent successfully completed the course in 12 out of 12 runs. Unlike conventional controllers, the trained agents demonstrated enhanced sailing capabilities like roll tacking and recovery from wind-stalled conditions. This work advances the understanding of autonomous sailing control and highlights pathways to bridge the reality gap, contributing to the broader adoption of RL in dynamic real-world applications.