Pith. sign in

Backtracking improves generation safety.arXiv preprint arXiv:2409.14586

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

fields

cs.LG 2

years

2026 2

verdicts

UNVERDICTED 2

representative citing papers

LLMs Should Express Uncertainty Explicitly

cs.LG · 2026-04-07 · unverdicted · novelty 6.0 · 2 refs

Training LLMs to verbalize uncertainty explicitly at the end or during reasoning reduces overconfident errors and improves answer quality on factual tasks while enabling RAG triggers.

citing papers explorer

Showing 2 of 2 citing papers.

  • Addressing Over-Refusal in LLMs with Competing Rewards cs.LG · 2026-06-30 · unverdicted · none · ref 77

    SEAR trains one LLM via adversarial process rewards to explore harmful reasoning paths but flip to safe outputs, reducing over-refusal while preserving safety.

  • LLMs Should Express Uncertainty Explicitly cs.LG · 2026-04-07 · unverdicted · none · ref 13 · 2 links

    Training LLMs to verbalize uncertainty explicitly at the end or during reasoning reduces overconfident errors and improves answer quality on factual tasks while enabling RAG triggers.