A hierarchical RL and meta-learning dialogue manager conditions an LLM for motivational interviewing and reports higher reward than a prompted LLM baseline in a simulated environment.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Tailored Conversations beyond LLMs: A RL-Based Dialogue Manager
A hierarchical RL and meta-learning dialogue manager conditions an LLM for motivational interviewing and reports higher reward than a prompted LLM baseline in a simulated environment.