Pith. sign in

REVIEW 2 cited by

Federated Deep Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1901.08277 v3 pith:NPXG26NO submitted 2019-01-24 cs.LG cs.AI

classification cs.LGcs.AI
keywords learningdeepmodelsreinforcementdataagentdomainsfederated
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In deep reinforcement learning, building policies of high-quality is challenging when the feature space of states is small and the training data is limited. Despite the success of previous transfer learning approaches in deep reinforcement learning, directly transferring data or models from an agent to another agent is often not allowed due to the privacy of data and/or models in many privacy-aware applications. In this paper, we propose a novel deep reinforcement learning framework to federatively build models of high-quality for agents with consideration of their privacies, namely Federated deep Reinforcement Learning (FedRL). To protect the privacy of data and models, we exploit Gausian differentials on the information shared with each other when updating their local models. In the experiment, we evaluate our FedRL framework in two diverse domains, Grid-world and Text2Action domains, by comparing to various baselines.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Federated Reinforcement Learning in Heterogeneous Environments

    cs.LG 2025-07 reject novelty 5.0 of 10

    FedRQ adds a robustness term to federated Q-learning and claims convergence to an optimal worst-case policy over heterogeneous local environments, but the proof has a reversed inequality.

  2. A Survey of Multi Agent Reinforcement Learning: Federated Learning and Cooperative and Noncooperative Decentralized Regimes

    cs.MA 2025-07 reject novelty 3.0 of 10

    A review of multi-agent reinforcement learning that catalogues federated, decentralized cooperative, and noncooperative regimes from the existing literature.

Pith tools