Pith. sign in

REVIEW 1 cited by

Action Semantics Network: Considering the Effects of Actions in Multiagent Systems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1907.11461 v3 pith:EWRAHKMY submitted 2019-07-26 cs.MA cs.AI

classification cs.MAcs.AI
keywords agentsactionsemanticsactionsmultiagentnetworkdifferentlearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In multiagent systems (MASs), each agent makes individual decisions but all of them contribute globally to the system evolution. Learning in MASs is difficult since each agent's selection of actions must take place in the presence of other co-learning agents. Moreover, the environmental stochasticity and uncertainties increase exponentially with the increase in the number of agents. Previous works borrow various multiagent coordination mechanisms into deep learning architecture to facilitate multiagent coordination. However, none of them explicitly consider action semantics between agents that different actions have different influences on other agents. In this paper, we propose a novel network architecture, named Action Semantics Network (ASN), that explicitly represents such action semantics between agents. ASN characterizes different actions' influence on other agents using neural networks based on the action semantics between them. ASN can be easily combined with existing deep reinforcement learning (DRL) algorithms to boost their performance. Experimental results on StarCraft II micromanagement and Neural MMO show ASN significantly improves the performance of state-of-the-art DRL approaches compared with several network architectures.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Novelty-Guided Data Reuse for Efficient and Diversified Multi-Agent Reinforcement Learning

    cs.LG 2024-12 conditional novelty 4.0 of 10

    MANGER uses RND-computed observation novelty to give each agent a different number of extra Q-learning updates, improving sample efficiency and behavioral diversity in cooperative MARL.

Pith tools