An LLM-based multi-agent framework successfully operates an AFM but benchmark results show current models are unreliable, instruction-following is fragile, and domain QA skill does not predict laboratory competence.
Title resolution pending
1 Pith paper cite this work, alongside 31 external citations. Polarity classification is still indexing.
1
Pith paper citing it
31
external citations · OpenAlex
citation-role summary
background 1
citation-polarity summary
fields
cs.CY 1years
2024 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Autonomous Microscopy Experiments through Large Language Model Agents
An LLM-based multi-agent framework successfully operates an AFM but benchmark results show current models are unreliable, instruction-following is fragile, and domain QA skill does not predict laboratory competence.