In a prostate radiotherapy simulator, iterative prompting of GPT-4V with Monte Carlo reward feedback produced better treatment plan scores than a DQN baseline and random gantry angles.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Transforming Multimodal Models into Action Models for Radiotherapy
In a prostate radiotherapy simulator, iterative prompting of GPT-4V with Monte Carlo reward feedback produced better treatment plan scores than a DQN baseline and random gantry angles.