Pith. sign in

Paper Citation Record · LEDGER

Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2403.18349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.18349 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:22:07.319424Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T15:38:56.326987Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 65fce4e2-dd40-4461-a55e-9a09d0b2d761 · inbound

Regression for the Mean: Auto-Evaluation and Inference with Few Labels through Post-hoc Regression cites this paper.

Regression for the Mean: Auto-Evaluation and Inference with Few Labels through Post-hoc Regression Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T17:22:07.319424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:22:07.319424Z digest=sha256:61cce4d75a17f01aa35d2dc0282b3bef182161c9d9d69a8a64454aadc15d77b4

Observation e2b5f1c6-d734-4c3a-837e-7738d504c989 · inbound

Reducing Tool Hallucination via Reliability Alignment cites this paper.

Reducing Tool Hallucination via Reliability Alignment Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.035095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.035095Z digest=sha256:50da2fde8557fe163a94c2c345b04dbebaa0ddbc125b1c67952a9caff8a05eb4

Observation a99bae8a-575c-4940-ac76-c709c416c0ae · inbound

GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation cites this paper.

GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T17:31:59.910590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:31:59.910590Z digest=sha256:7d67bb055070ef4922d4edb26c39814af2d89a1ca28d04409765bd558234a19f

Observation 702d2ad4-02d6-4f69-82be-ff06d645237e · inbound

Security Concerns for Large Language Models: A Survey cites this paper.

Security Concerns for Large Language Models: A Survey Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:27:03.720278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:27:03.720278Z digest=sha256:7fe8bf9ff512c89044ba6246c763144568884e7da74d9c74aea58991cfb4cf48

Observation 3509225c-0ef7-4b5f-a713-76e4a1114a3e · inbound

Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing cites this paper.

Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:24.985286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:24.985286Z digest=sha256:0aff2b440588848490b075546ef8de4a1663d5542d56e5ab41100709ebfe8e8f

Observation cb0e920c-de7e-41a8-8e2a-6c3551cf4553 · inbound

Building Task Bots with Self-learning for Enhanced Adaptability, Extensibility, and Factuality cites this paper.

Building Task Bots with Self-learning for Enhanced Adaptability, Extensibility, and Factuality Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

Reference 201

Resolution
verified exact
local_arxiv, observed 2026-08-05T15:38:56.332107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-05T15:38:55.221247Z digest=sha256:d7cd680d5d3aaf67d65d66db6e9153dfe18dd2a804e6ecd27f516cb631ac0e63

Observation 74247269-072c-4a8b-a849-8dbaae796dad · inbound

Abstention as an Action Can Kill Both the Reward Gradient and the KL Anchor: Collapse Law and Repair for Error-Penalized Reinforcement Learning cites this paper.

Abstention as an Action Can Kill Both the Reward Gradient and the KL Anchor: Collapse Law and Repair for Error-Penalized Reinforcement Learning Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T00:53:26.819594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:53:26.819594Z digest=sha256:09efb8e9fccd383fe8401cc12aa2d7a27c0cce6d6f808e22f9f00d89dd55f48b