REVIEW 5 cited by
Quantum Advantage Actor-Critic for Reinforcement Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Quantum computing offers efficient encapsulation of high-dimensional states. In this work, we propose a novel quantum reinforcement learning approach that combines the Advantage Actor-Critic algorithm with variational quantum circuits by substituting parts of the classical components. This approach addresses reinforcement learning's scalability concerns while maintaining high performance. We empirically test multiple quantum Advantage Actor-Critic configurations with the well known Cart Pole environment to evaluate our approach in control tasks with continuous state spaces. Our results indicate that the hybrid strategy of using either a quantum actor or quantum critic with classical post-processing yields a substantial performance increase compared to pure classical and pure quantum variants with similar parameter counts. They further reveal the limits of current quantum approaches due to the hardware constraints of noisy intermediate-scale quantum computers, suggesting further research to scale hybrid approaches for larger and more complex control tasks.
Forward citations
Cited by 5 Pith papers
-
Quantum entanglement provides a competitive advantage in adversarial games
Entangled 8-qubit PQC feature extractors in PPO agents for Pong consistently beat separable PQCs of similar size and can match or exceed small classical MLPs in the low-parameter regime.
-
PPO-Q: Proximal Policy Optimization with Parametrized Quantum Policies or Values
PPO-Q combines a small parameterized quantum circuit with pre-encoding and post-processing neural networks inside the PPO algorithm, matching classical performance on eight tasks with fewer parameters and solving Bipe...
-
Quantum Reinforcement Learning by Adaptive Non-local Observables
Adaptive non-local observables, jointly trained with variational circuit parameters, improve DQN and A3C reinforcement learning agents on simulated benchmark tasks relative to fixed Pauli-measurement baselines.
-
Quantum computing and artificial intelligence: status and perspectives
A broad expert white paper sets a European research agenda for combining quantum computing and AI, spanning quantum machine learning, AI-driven quantum control, and foundational questions.
-
Introduction to Quantum Machine Learning and Quantum Architecture Search
A tutorial reviewing quantum machine learning models and automated quantum architecture search methods, with no new experiments or derivations.
Discussion (0). Continue with ORCID to comment.