Pith. sign in

REVIEW 1 cited by

Variational Quantum Circuit Design for Quantum Reinforcement Learning on Continuous Environments

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2312.13798 v1 pith:A2JC67CT submitted 2023-12-21 quant-ph

classification quant-ph
keywords designquantumclassicalenvironmentslearningreinforcementactionagents
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Quantum Reinforcement Learning (QRL) emerged as a branch of reinforcement learning (RL) that uses quantum submodules in the architecture of the algorithm. One branch of QRL focuses on the replacement of neural networks (NN) by variational quantum circuits (VQC) as function approximators. Initial works have shown promising results on classical environments with discrete action spaces, but many of the proposed architectural design choices of the VQC lack a detailed investigation. Hence, in this work we investigate the impact of VQC design choices such as angle embedding, encoding block architecture and postprocessesing on the training capabilities of QRL agents. We show that VQC design greatly influences training performance and heuristically derive enhancements for the analyzed components. Additionally, we show how to design a QRL agent in order to solve classical environments with continuous action spaces and benchmark our agents against classical feed-forward NNs.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Benchmarking Quantum Reinforcement Learning

    quant-ph 2025-02 conditional novelty 6.0 of 10

    A gridworld benchmark shows amplitude-amplification QRL outperforms PQC and free-energy QRL on cost and clock time, while entanglement and replica-count ablations find little evidence that current QRL performance depe...

Pith tools