A benchmark of 49 QHack PennyLane challenges shows LLMs solve at most 49 percent of tasks, retrieval augmentation usually does not help, and a multi-agent retry loop improves results.
Demonstrating quantum advantage in hybrid quantum neural networks for model capacity,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
QHackBench: Benchmarking Large Language Models for Quantum Code Generation Using PennyLane Hackathon Challenges
A benchmark of 49 QHack PennyLane challenges shows LLMs solve at most 49 percent of tasks, retrieval augmentation usually does not help, and a multi-agent retry loop improves results.