Pith. sign in

REVIEW

PAC Guarantees for Cooperative Multi-Agent Reinforcement Learning with Restricted Communication

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1905.09951 v2 pith:XRFZDJSX submitted 2019-05-23 cs.LG stat.ML

classification cs.LGstat.ML
keywords communicationagentsguaranteesnoisefreeimprovedinformationlimited
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We develop model free PAC performance guarantees for multiple concurrent MDPs, extending recent works where a single learner interacts with multiple non-interacting agents in a noise free environment. Our framework allows noisy and resource limited communication between agents, and develops novel PAC guarantees in this extended setting. By allowing communication between the agents themselves, we suggest improved PAC-exploration algorithms that can overcome the communication noise and lead to improved sample complexity bounds. We provide a theoretically motivated algorithm that optimally combines information from the resource limited agents, thereby analyzing the interaction between noise and communication constraints that are ubiquitous in real-world systems. We present empirical results for a simple task that supports our theoretical formulations and improve upon naive information fusion methods.

Discussion (0). Continue with ORCID to comment.

Pith tools