Pith. sign in

REVIEW 1 cited by

Policy Diagnosis via Measuring Role Diversity in Cooperative Multi-agent RL

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2207.05683 v1 pith:TFNU37AI submitted 2022-06-01 cs.MA cs.AIcs.LG

classification cs.MAcs.AIcs.LG
keywords multi-agentdiversitypolicyrolemarltaskthreebehavior
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Cooperative multi-agent reinforcement learning (MARL) is making rapid progress for solving tasks in a grid world and real-world scenarios, in which agents are given different attributes and goals, resulting in different behavior through the whole multi-agent task. In this study, we quantify the agent's behavior difference and build its relationship with the policy performance via {\bf Role Diversity}, a metric to measure the characteristics of MARL tasks. We define role diversity from three perspectives: action-based, trajectory-based, and contribution-based to fully measure a multi-agent task. Through theoretical analysis, we find that the error bound in MARL can be decomposed into three parts that have a strong relation to the role diversity. The decomposed factors can significantly impact policy optimization on three popular directions including parameter sharing, communication mechanism, and credit assignment. The main experimental platforms are based on {\bf Multiagent Particle Environment (MPE)} and {\bf The StarCraft Multi-Agent Challenge (SMAC). Extensive experiments} clearly show that role diversity can serve as a robust measurement for the characteristics of a multi-agent cooperation task and help diagnose whether the policy fits the current multi-agent system for a better policy performance.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Novelty-Guided Data Reuse for Efficient and Diversified Multi-Agent Reinforcement Learning

    cs.LG 2024-12 conditional novelty 4.0 of 10

    MANGER uses RND-computed observation novelty to give each agent a different number of extra Q-learning updates, improving sample efficiency and behavioral diversity in cooperative MARL.

Pith tools