LLM yes-no bias on moral dilemmas is an order-plus-lexical surface artifact, not a moral shift; models have a nearly format-invariant graded stance that the standard binary readout confounds.
Title resolution pending
1 Pith paper cite this work, alongside 462 external citations. Polarity classification is still indexing.
1
Pith paper citing it
462
external citations · OpenAlex
fields
cs.CL 1years
2026 1verdicts
ACCEPT 1representative citing papers
citing papers explorer
-
The yes-no bias of large language models reflects answer order and wording, not shifts in moral judgment
LLM yes-no bias on moral dilemmas is an order-plus-lexical surface artifact, not a moral shift; models have a nearly format-invariant graded stance that the standard binary readout confounds.