Pith. sign in

REVIEW 1 cited by

Deep Reinforcement Learning with Decorrelation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1903.07765 v3 pith:XSWSRNMH submitted 2019-03-18 cs.LG cs.AIstat.ML

classification cs.LGcs.AIstat.ML
keywords learninggamesalgorithmsdeepdecorrelationfeaturesnetworksreinforcement
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Learning an effective representation for high-dimensional data is a challenging problem in reinforcement learning (RL). Deep reinforcement learning (DRL) such as Deep Q networks (DQN) achieves remarkable success in computer games by learning deeply encoded representation from convolution networks. In this paper, we propose a simple yet very effective method for representation learning with DRL algorithms. Our key insight is that features learned by DRL algorithms are highly correlated, which interferes with learning. By adding a regularized loss that penalizes correlation in latent features (with only slight computation), we decorrelate features represented by deep neural networks incrementally. On 49 Atari games, with the same regularization factor, our decorrelation algorithms perform $70\%$ in terms of human-normalized scores, which is $40\%$ better than DQN. In particular, ours performs better than DQN on 39 games with 4 close ties and lost only slightly on $6$ games. Empirical results also show that the decorrelation method applies to Quantile Regression DQN (QR-DQN) and significantly boosts performance. Further experiments on the losing games show that our decorelation algorithms can win over DQN and QR-DQN with a fined tuned regularization factor.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning

    cs.LG 2025-01 conditional novelty 5.0 of 10

    Decorrelated Soft Actor-Critic (DSAC) adds layerwise input decorrelation to discrete SAC and reports faster wall-clock training in 5 of 7 Atari games and better reward in 2, though the gains are partly confounded by p...

Pith tools