NeFut Logo NeFut
Admin Login

[CS.AI] Reinforcement Learning Framework for Heterogeneous Sensor Selection in Maritime Surveillance

Published at: 2026-07-29 22:00 Last updated: 2026-07-30 03:24
#Reinforcement Learning #Sensor Selection #Maritime Surveillance

This paper presents an information-gain-guided reinforcement-learning sensor-selection framework for single-vessel tracking in heterogeneous maritime sensor networks. The proposed approach is motivated by information-theoretic sensor management: rather than activating all sensors or repeatedly performing computationally expensive online expected-information-gain evaluation, a learned policy selects one tracking-relevant sensor at each decision epoch.\n\nA Bayesian sequential Monte Carlo tracker estimates the vessel state from noisy measurements and provides a belief representation for scheduling under nonlinear and non-Gaussian conditions. A Proximal Policy Optimization agent selects one of five sensors deployed in a georeferenced simulation of the CMMI Smart Marina testbed at Ayia Napa Marina, Cyprus. The agent observes belief-state, detection-history, coverage, sensor-geometry, and realized-information-gain features. The reward is defined as a realized-information-gain term gated by an observability mask.\n\nFinal-test simulations compare the proposed framework with random single-sensor selection, always-on sensing using all sensors simultaneously, and the expected-information-gain sensor-selection baseline proposed in our previous work. Results show that the learned policy achieves tracking performance close to always-on sensing while activating only one sensor per decision time step and avoiding the computationally expensive online entropy search required by expected-information-gain selection.\n\nBlogger's Review: This paper demonstrates the potential of optimizing sensor selection through reinforcement learning, effectively utilizing information gain in complex environments. The proposed framework strikes a good balance between performance and computational efficiency compared to traditional methods, offering significant application prospects. The implementation details and experimental results provide a solid foundation for further research.

Original Source: https://arxiv.org/abs/2607.22667

[h] Back to Home