An AlphaZero-like RL agent can synthesize exact Clifford+T circuits for up to three qubits with ancilla, recovering known optimal decompositions and a 4-T Toffoli implementation.
Towards "AlphaChem": Chemical Synthesis Planning with Tree Search and Deep Neural Network Policies
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Retrosynthesis is a technique to plan the chemical synthesis of organic molecules, for example drugs, agro- and fine chemicals. In retrosynthesis, a search tree is built by analysing molecules recursively and dissecting them into simpler molecular building blocks until one obtains a set of known building blocks. The search space is intractably large, and it is difficult to determine the value of retrosynthetic positions. Here, we propose to model retrosynthesis as a Markov Decision Process. In combination with a Deep Neural Network policy learned from essentially the complete published knowledge of chemistry, Monte Carlo Tree Search (MCTS) can be used to evaluate positions. In exploratory studies, we demonstrate that MCTS with neural network policies outperforms the traditionally used best-first search with hand-coded heuristics.
fields
quant-ph 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Unitary Synthesis with AlphaZero via Dynamic Circuits
An AlphaZero-like RL agent can synthesize exact Clifford+T circuits for up to three qubits with ancilla, recovering known optimal decompositions and a 4-T Toffoli implementation.