RL-Based Sensor Selection for Maritime Surveillance
July 28, 2026
A Proximal Policy Optimization agent selects from five heterogeneous sensors to optimize tracking in maritime environments. The framework uses a Bayesian sequential Monte Carlo tracker to manage belief states under non-Gaussian conditions, reducing the computational overhead of online information-gain evaluations.
HOW THIS AFFECTS YOU
●
researcherYou can apply this information-theoretic RL approach to multi-sensor scheduling problems in non-linear environments.