A JAX-based GPU environment and a transformer-plus-curriculum MARL method train policies that transfer to the Gazebo LRAUV simulator and track up to 5 targets with around 5 m average error.
Frontal dynamics in the Alboran sea: 1. Coherent 3D pathways at the Almeria- Oran front using underwater glider observations
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.RO 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Scaling Multi Agent Reinforcement Learning for Underwater Acoustic Tracking via Autonomous Vehicles
A JAX-based GPU environment and a transformer-plus-curriculum MARL method train policies that transfer to the Gazebo LRAUV simulator and track up to 5 targets with around 5 m average error.