Pith. sign in

REVIEW 1 cited by

Robustness to Multi-Modal Environment Uncertainty in MARL using Curriculum Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.08746 v1 pith:7DGPR6DZ submitted 2023-10-12 cs.LG

Robustness to Multi-Modal Environment Uncertainty in MARL using Curriculum Learning

classification cs.LG
keywords uncertaintyenvironmentmarllearningmulti-modalreal-worldrobustnessapproach
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

Multi-agent reinforcement learning (MARL) plays a pivotal role in tackling real-world challenges. However, the seamless transition of trained policies from simulations to real-world requires it to be robust to various environmental uncertainties. Existing works focus on finding Nash Equilibrium or the optimal policy under uncertainty in one environment variable (i.e. action, state or reward). This is because a multi-agent system itself is highly complex and unstationary. However, in real-world situation uncertainty can occur in multiple environment variables simultaneously. This work is the first to formulate the generalised problem of robustness to multi-modal environment uncertainty in MARL. To this end, we propose a general robust training approach for multi-modal uncertainty based on curriculum learning techniques. We handle two distinct environmental uncertainty simultaneously and present extensive results across both cooperative and competitive MARL environments, demonstrating that our approach achieves state-of-the-art levels of robustness.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations

    cs.CV 2025-09 conditional novelty 6.0

    RobustVLA benchmarks VLA robot policies under 17 multi-modal perturbations and uses adversarial flow-matching plus UCB-based noise selection to raise robustness by up to 12.6 absolute points on LIBERO.