著者
Takefumi Higaki Hirotada Hashimoto Hitoshi Yoshioka
出版者
公益社団法人 日本船舶海洋工学会
雑誌
日本船舶海洋工学会論文集 (ISSN:18803717)
巻号頁・発行日
vol.36, pp.137-148, 2022 (Released:2023-01-31)
参考文献数
32

Automatic collision avoidance is of significant importance to prevent maritime collisions. Although many studies have been conducted in recent years, autonomous system has not completely replaced human captains since it is still difficult to imitate their complicated decisions. Thus, the present paper tries to investigate and imitate experienced captains' maneuver using maximum entropy inverse reinforcement learning (MaxEnt IRL). We firstly verify that MaxEnt IRL can reproduce appropriate reward function from demonstrative trajectories. Afterwards, we conduct an experiment on a simulator where well-experienced captains maneuver in congested sea and estimate reward from the trajectories. Searching the route which maximizes the obtained reward, finally, we demonstrate the optimized route can avoid collision against multiple ships in compliance with the International Regulations for Preventing Collisions at Sea (COLREGs).