Interactive LLM dialogue raised residents' hard-case diagnostic correctness from 0.589 to 0.734 and produced medium effect sizes in a blinded study of seven physicians on 52 emergency cases.
A Demonstration of Adaptive Collaboration of Large Language Models for Medical Decision-Making
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
fields
cs.AI 2years
2026 2verdicts
UNVERDICTED 2representative citing papers
Second-order dynamical integration of LLM risk outputs produces smooth anticipatory concern trajectories in synthetic ward scenarios, unlike the sharp cliffs from stateless agents.
citing papers explorer
-
Human-LLM Dialogue Improves Diagnostic Accuracy in Emergency Care
Interactive LLM dialogue raised residents' hard-case diagnostic correctness from 0.589 to 0.734 and produced medium effect sizes in a blinded study of seven physicians on 52 emergency cases.
-
Modeling Clinical Concern Trajectories in Language Model Agents
Second-order dynamical integration of LLM risk outputs produces smooth anticipatory concern trajectories in synthetic ward scenarios, unlike the sharp cliffs from stateless agents.