Pith. sign in

Paper Citation Record · LEDGER

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents

As of 7 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2607.19449.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.19449 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T13:32:29.381847Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 18dd4466-8d49-4712-b031-27e77478d38b · outbound

This paper cites an unresolved cited work.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.256476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.256476Z digest=sha256:bc930260e933f105de098f65b5d936fe8265296992985628155690a7e2f1454b

Observation bd1b231e-7fde-42ba-9674-4e06797bdbb8 · outbound

This paper cites 2022.LangChain.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents 2022.LangChain

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.412560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.412560Z digest=sha256:a0e45015b651831f6458316b2f919cffe3c4cf2466edb130509ef876e7f454d4

Observation 6ad12b20-9788-438a-8cf0-e574c99615c3 · outbound

This paper cites The Llama 3 Herd of Models.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.555493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.555493Z digest=sha256:c1c8e4953060d88f1f388b756096f99d31055fb152974b0d367a395f0ed7c81a

Observation 7dcb1b53-9b6e-46c7-8517-7893b3e65c5a · outbound

This paper cites Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.638390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.638390Z digest=sha256:cbb8dece8338cfd95f5cfac33d83dc798998fe31efd18fba6926b160d88ba9fc

Observation 97fb5f10-7320-4508-911d-19fc2702a96a · outbound

This paper cites an unresolved cited work.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.698688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.698688Z digest=sha256:a16f2a9d8c40f1c9beb8ce38446fb4cc44ff48ced679764860a5d58b43261db6

Observation 4e784f80-2647-493d-8b47-f2a4b97f77fa · outbound

This paper cites an unresolved cited work.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.754948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.754948Z digest=sha256:dd8fc0b4a070aa5bee4def98ea6b7e84de569634c8bd3fb10ece8fca20742ac4

Observation 1e0e3b43-9f07-43af-828d-bfca0611122a · outbound

This paper cites AgentBench: Evaluating LLMs as Agents.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents AgentBench: Evaluating LLMs as Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.851035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.851035Z digest=sha256:a6c261889d085016fe3dc24ad983bfaea1c78ba663add4b3eb5e00b3325428d7

Observation fcb696dc-df53-466f-bb80-3cafc00c1e03 · outbound

This paper cites GPT-4 Technical Report.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents GPT-4 Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.941780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.941780Z digest=sha256:d043d267ae6a6ead2a40182e1f8d1236294595aea0a8550cdc71e28b9704e4be

Observation 6df75736-952d-45a1-a109-65ce7453fbf8 · outbound

This paper cites Toolformer: Language Models Can Teach Themselves to Use Tools.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Toolformer: Language Models Can Teach Themselves to Use Tools

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.984369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.984369Z digest=sha256:c7ee92c99f8b4bd8c3facde0118866ac43080eaf17dc9290c4fb0f482ad784a5

Observation 5d372112-1359-407f-a96a-bbd85947bd62 · outbound

This paper cites Towards Understanding Sycophancy in Language Models.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Towards Understanding Sycophancy in Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.021850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.021850Z digest=sha256:dc4ff7bf83a51d6ccee8db680b0162ddbfbcc79bbdde12b24ebf66028d7a1d9e

Observation 7198e472-2e76-4e26-8d90-da064b4f76cc · outbound

This paper cites Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.076683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.076683Z digest=sha256:187d667103c8008284903f2e5deddac4b4d02126731ba0938d538b6b46968937

Observation 1afc3dad-63fd-47c2-b1e8-46643423f7e0 · outbound

This paper cites an unresolved cited work.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.116487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.116487Z digest=sha256:1c4bb8ba07ffc6cfc866fd97027b86609308a864868e7b1d884b0cbba9b46ec9

Observation 1672d1f7-f840-4c5a-88d9-0909195994ad · outbound

This paper cites SafeToolBench: Pioneering a Prospective Benchmark to Evaluating Tool Utilization Safety in LLMs.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents SafeToolBench: Pioneering a Prospective Benchmark to Evaluating Tool Utilization Safety in LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.176874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.176874Z digest=sha256:bc60adceb6d8b9d58a0209096aeaabc7bcc3febce66ce2ccdcbabc79f0b8796d

Observation 75d6808a-4b51-4bb6-a7fb-428db49446be · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents ReAct: Synergizing Reasoning and Acting in Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.283897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.283897Z digest=sha256:3dee4b27fde1f83d002b2f9a76c65c50069b5c573d14b7402035b0ec9b4b1e05

Observation 864a4ce6-ffbc-4913-b6a3-d864a99c8d07 · outbound

This paper cites the tool returned no data.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents the tool returned no data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.381847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.381847Z digest=sha256:2edf1e313f660c7a058897be34de80b275bd1ccf51e0f5196ca7f714107179be

Pith citing papers

No inbound Pith citation observations are available.