Pith. sign in

Paper Citation Record · LEDGER

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents

As of 23 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2607.19449.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.19449 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T13:32:29.381847Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 18dd4466-8d49-4712-b031-27e77478d38b · outbound

This paper cites an unresolved cited work.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.256476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.256476Z digest=sha256:037d2f68b4239608ef3a582a7a408b6bc0469179a26aa2fa69bc7ab2eb5918a8

Observation bd1b231e-7fde-42ba-9674-4e06797bdbb8 · outbound

This paper cites 2022.LangChain.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents 2022.LangChain

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.412560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.412560Z digest=sha256:4f135fe708deebbb2a23ceb1822e09251b8a95cca575f536f30050d1ac9af59d

Observation 6ad12b20-9788-438a-8cf0-e574c99615c3 · outbound

This paper cites The Llama 3 Herd of Models.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.555493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.555493Z digest=sha256:c25a63ccbe8a95082e4a217ec6115e15b1e5ed0826e61b4208c7d7a0604c4ba9

Observation 7dcb1b53-9b6e-46c7-8517-7893b3e65c5a · outbound

This paper cites Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.638390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.638390Z digest=sha256:c178590fb87a8c8c07f4f280c13db82b4a50b98825da4e5f1702bc8292841624

Observation 97fb5f10-7320-4508-911d-19fc2702a96a · outbound

This paper cites an unresolved cited work.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.698688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.698688Z digest=sha256:a6147491bb96bb2e3c61f40e0e7c16ac0bf22693d7633aececda9d1123bb4de7

Observation 4e784f80-2647-493d-8b47-f2a4b97f77fa · outbound

This paper cites an unresolved cited work.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.754948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.754948Z digest=sha256:216f1311f9bc8cca3eab389ab73a90cd49981965791d476797618fe40187dbb3

Observation 1e0e3b43-9f07-43af-828d-bfca0611122a · outbound

This paper cites AgentBench: Evaluating LLMs as Agents.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents AgentBench: Evaluating LLMs as Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.851035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.851035Z digest=sha256:3c5def064dba71c11227d3d8dc6ed78b5a27115b899f862f910e8820e11e1f49

Observation fcb696dc-df53-466f-bb80-3cafc00c1e03 · outbound

This paper cites GPT-4 Technical Report.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents GPT-4 Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.941780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.941780Z digest=sha256:866a342d1e3e498edf3a7b86bf3a1e18d9ed038c3dc6e14f5766a80e5bdb56c9

Observation 6df75736-952d-45a1-a109-65ce7453fbf8 · outbound

This paper cites Toolformer: Language Models Can Teach Themselves to Use Tools.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Toolformer: Language Models Can Teach Themselves to Use Tools

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:28.984369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:28.984369Z digest=sha256:07ae939061545515aed367cc198742f60a2d36b0bc7b283be69f53dbd1ed9910

Observation 5d372112-1359-407f-a96a-bbd85947bd62 · outbound

This paper cites Towards Understanding Sycophancy in Language Models.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Towards Understanding Sycophancy in Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.021850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.021850Z digest=sha256:3f6ea823373e6a2c9a8ebb7f7aa07ebcac2a37e856f151a4ac9400970833703e

Observation 7198e472-2e76-4e26-8d90-da064b4f76cc · outbound

This paper cites Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.076683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.076683Z digest=sha256:db5f7dbb5cfa3b1f824fd8c34830827327bb5b96d54ad84e7fcac388ceb36e43

Observation 1afc3dad-63fd-47c2-b1e8-46643423f7e0 · outbound

This paper cites an unresolved cited work.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.116487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.116487Z digest=sha256:ce236d522232fe91f2da0f851a4d4684d855ea8720bcc7e04ea45063d2097a35

Observation 1672d1f7-f840-4c5a-88d9-0909195994ad · outbound

This paper cites SafeToolBench: Pioneering a Prospective Benchmark to Evaluating Tool Utilization Safety in LLMs.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents SafeToolBench: Pioneering a Prospective Benchmark to Evaluating Tool Utilization Safety in LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.176874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.176874Z digest=sha256:8acd69798d06be0fabc331c005c9ac5739c0e510ef3267a5889f827b942bef64

Observation 75d6808a-4b51-4bb6-a7fb-428db49446be · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents ReAct: Synergizing Reasoning and Acting in Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.283897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.283897Z digest=sha256:efb83b985ed6170045a9276d5d14d9d77a454622e9978fe2f06db8938a27edb1

Observation 864a4ce6-ffbc-4913-b6a3-d864a99c8d07 · outbound

This paper cites the tool returned no data.

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents the tool returned no data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T13:32:29.381847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:32:29.381847Z digest=sha256:aae41c8d373b3fda19e410198efe993ab49e10f1ced7ecff59b10b24e03c66df

Pith citing papers

No inbound Pith citation observations are available.