A JAX-based GPU environment and a transformer-plus-curriculum MARL method train policies that transfer to the Gazebo LRAUV simulator and track up to 5 targets with around 5 m average error.
A system of coordinated au- tonomous robots for Lagrangian studies of microbes in the oceanic deep chlorophyll maximum
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.RO 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Scaling Multi Agent Reinforcement Learning for Underwater Acoustic Tracking via Autonomous Vehicles
A JAX-based GPU environment and a transformer-plus-curriculum MARL method train policies that transfer to the Gazebo LRAUV simulator and track up to 5 targets with around 5 m average error.