Pith. sign in

Paper Citation Record · LEDGER

ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2402.11889.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.11889 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T16:18:40.699323Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:29:44.263692Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 123dee81-6968-497b-8966-f3b1c026d33a · inbound

On Almost Surely Safe Alignment of Large Language Models at Inference-Time cites this paper.

On Almost Surely Safe Alignment of Large Language Models at Inference-Time ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T16:18:40.699323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:18:40.699323Z digest=sha256:26244e5d1bc989b876fb11033cd6eb14df82cd26326ea33d1238d9d78a0256a5

Observation 1253af54-4495-4576-919c-9a0b5f21ff5b · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 214

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.324707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:f55f10d8d17ba40ccef1d167de236f11804a831c5d740379ec1495258e0515c2

Observation 2ed13e14-0892-4b18-a2a6-46c16f8511df · inbound

Exploring and Mitigating Fawning Hallucinations in Large Language Models cites this paper.

Exploring and Mitigating Fawning Hallucinations in Large Language Models ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T13:11:15.385658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:11:15.385658Z digest=sha256:29bf4816212d089e761c8d755f2e78686e0a574d849d5eef03cb1acac223b762

Observation 03628607-c82e-43f1-b156-b6511d09dde8 · inbound

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs cites this paper.

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:29:44.265560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T09:52:52.603708Z digest=sha256:838dfd16473dc68abbd11d8583c39fd823b853a2518e3bb3223d981d90f3eecc

Observation 7d9809cf-d8ef-40c9-a3a9-8b10e5f846e8 · inbound

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs cites this paper.

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:55:35.054465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T07:07:45.177242Z digest=sha256:8efc7919b10f274dd3b356fa6e2443aa1a201827d978db046ae73efed99fa900