Pith. sign in

Paper Citation Record · LEDGER

Should LLM Safety Be More Than Refusing Harmful Instructions?

As of 7 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2506.02442.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02442 v2

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:27:47.482807Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact3
  • verified fuzzy2
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4f626177-35eb-4ab9-bd5b-a0c7274a307b · outbound

This paper cites GPT-4 Technical Report.

Should LLM Safety Be More Than Refusing Harmful Instructions? GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.122877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.122877Z digest=sha256:95f2748c883f8b8f2c30ca361c78dfd4c489c0b31c9073198cc922a8aebb202f

Observation 52404c4d-5d6a-471e-a625-cda4dadbb19f · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:50.178348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.129536Z digest=sha256:a96efedba868e895d89f5a7a472e3cc9aef68eff04d9da1a64b5daf1587f3d12

Observation ae55402e-b8c8-4b91-8a4b-b978302c4afb · outbound

This paper cites Detecting Language Model Attacks with Perplexity.

Should LLM Safety Be More Than Refusing Harmful Instructions? Detecting Language Model Attacks with Perplexity

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.137122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.137122Z digest=sha256:92a688411ea0b6af93f38ea2bbec34033557076877c1c035f5b2c4b6c3c5bd65

Observation 04624ff3-35b9-404b-ab65-c8b96582c804 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? Gemini: A Family of Highly Capable Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.144854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.144854Z digest=sha256:5637eb93bf795ca936c99253bb5c080b65b23052105034856c4364f89a44133e

Observation cc9f277e-c466-46c5-ac76-a412e39331ef · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:50.095009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.152801Z digest=sha256:394d9aa4b895d5f05013f56f2fc862623b6816f321f24dd32077eec45d47dcb1

Observation a2c1901f-66e2-4c83-a48e-6db1d4fb6c5c · outbound

This paper cites Jailbreaking Large Language Models with Symbolic Mathematics.

Should LLM Safety Be More Than Refusing Harmful Instructions? Jailbreaking Large Language Models with Symbolic Mathematics

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.159009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.159009Z digest=sha256:46ace10d1959984d3a98e07f38e125fcd1dc0573ab21b8337087903474c24120

Observation f11eae64-55f7-45ac-b29d-bf93198671f2 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.166556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.166556Z digest=sha256:ccf54aed8ca87b74a3c8db82020c4eb1e11daf76722c18ebf307acb7edc6aac2

Observation bbb14b9f-4929-4411-a53d-120cf8f8b928 · outbound

This paper cites Pappas, Florian Tram\` e r, Hamed Hassani, and Eric Wong.

Should LLM Safety Be More Than Refusing Harmful Instructions? Pappas, Florian Tram\` e r, Hamed Hassani, and Eric Wong

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:27:49.928300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.175424Z digest=sha256:c5d9cad57276d50f4b50032d1328625c18476eee1a27c93d64e3a8e6f22a7199

Observation c43bbd36-f338-4770-bffb-980105f6b432 · outbound

This paper cites Pappas, Florian Tram \`e r, Hamed Hassani, and Eric Wong.

Should LLM Safety Be More Than Refusing Harmful Instructions? Pappas, Florian Tram \`e r, Hamed Hassani, and Eric Wong

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:27:49.777409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.182198Z digest=sha256:311d7779d35d2b32e7b5704fc3ae84ef8c717bcdde5721af99cccaaf47a85d1e

Observation 3ecfed4b-90af-4563-bbef-b163fb4546b3 · outbound

This paper cites Recent Advances in Attack and Defense Approaches of Large Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? Recent Advances in Attack and Defense Approaches of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.188977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.188977Z digest=sha256:4ae4967cfcab2539f8b4a7b8ac90c269dceebab19645c29bfc8c5a48d34eb266

Observation 5ca62124-816a-4f1f-9737-e4b7ec72e693 · outbound

This paper cites Attacks, Defenses and Evaluations for LLM Conversation Safety: A Survey.

Should LLM Safety Be More Than Refusing Harmful Instructions? Attacks, Defenses and Evaluations for LLM Conversation Safety: A Survey

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.195714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.195714Z digest=sha256:3f10e46d5b1121cbf480b99758b4aeb37009c5bd6ec1c8da47750ea6711b259a

Observation c6b7078e-91a8-422d-af32-06c9c369cce8 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 12

Resolution
verified exact
doi, observed 2026-08-07T11:27:47.649065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.204744Z digest=sha256:e37f1cb49923866c2b2dec184e06ba2445c19be7835ebbe0aa44ed1ef5dff154

Observation 72ae24e7-ab71-481b-9ca2-3158ebd550e2 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.647650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.213100Z digest=sha256:618c77dc114a3ac3a444eb5779a88a16f71d6612020ca0687b5e635a20a6d9b8

Observation 88b1e8b6-d140-46b4-a892-d0a8b718f7a3 · outbound

This paper cites Unsupervised Cipher Cracking Using Discrete GANs.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unsupervised Cipher Cracking Using Discrete GANs

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:27:48.698390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.227011Z digest=sha256:e4d77cd9d44f8fd6f98e195ac1d3c7ec2c036cd24ac53c620e14b464c7d17c3f

Observation 30ab4d27-9b68-4ebd-98b4-569097c9ddc8 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Should LLM Safety Be More Than Refusing Harmful Instructions? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.235852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.235852Z digest=sha256:b386760157714ae99a1b3b36b549ea04517d7cccb14cf3bc6e8563eba96138ed

Observation 567d72f0-94fd-42bf-9c2d-ba662c14ee19 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.501436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.242254Z digest=sha256:78aa2bc704923fa37bdfdb27b7ec598c24fc017f6ace4292a2e1dca510d46aa0

Observation b07cb090-4083-40a9-ae2b-5094165514ae · outbound

This paper cites competency.

Should LLM Safety Be More Than Refusing Harmful Instructions? competency

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.250751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.250751Z digest=sha256:94b1ad78b2738a05bb524dd1efe59f8e40f905b4177ed5f25e2489676a5dff25

Observation f21550f7-2f32-4419-af9e-5c182045b666 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.259907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.259907Z digest=sha256:38874080d626668088329dc09cdca0ead3d3717ab4ab4f3d431ec1ee661e5fba

Observation 5444761c-5f45-403b-93b7-032d835e6c19 · outbound

This paper cites Endless Jailbreaks with Bijection Learning.

Should LLM Safety Be More Than Refusing Harmful Instructions? Endless Jailbreaks with Bijection Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.272740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.272740Z digest=sha256:e22e1f18d9f886e360e8a856889cfc33172e722705dd7fef5f14972daeb00cc1

Observation 88b28cec-7931-4bd0-ab13-fd3bf806ad74 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Should LLM Safety Be More Than Refusing Harmful Instructions? Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.280600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.280600Z digest=sha256:c1e1712d28b967077967d3d36a894111fbe1c3d69783afe81c9719d6482d9e90

Observation d449b9ff-6312-494c-9fce-64e26117404f · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.286452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.286452Z digest=sha256:c756f1197801b8582cf553572374235db4c9ec48f827d374c3408ea78895e889

Observation 73344d01-9c0e-4c79-b7c9-733d12031035 · outbound

This paper cites Mistral 7B.

Should LLM Safety Be More Than Refusing Harmful Instructions? Mistral 7B

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.295083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.295083Z digest=sha256:d059954b5f4a50a3e940c88bffa9c8d62af2c3c8d032978092a914a816e66925

Observation 7093f919-ff68-48f4-8443-7131b57c8455 · outbound

This paper cites ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs.

Should LLM Safety Be More Than Refusing Harmful Instructions? ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.302318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.302318Z digest=sha256:eb0d718273b18bf08ec3bad47949b945e1f8c749adfd1dab8fddaf1ee1755116

Observation d5aaf8e2-e118-45dc-a18d-0d865ef72af0 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.310189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.310189Z digest=sha256:1a0f068811419a538c73bbaa5bf2a1ead8020e1184bae377f401bbd8be69863e

Observation 25be7c48-77af-4c52-8fb1-e932c0bf2cd7 · outbound

This paper cites CipherBank: Exploring the Boundary of LLM Reasoning Capabilities through Cryptography Challenges.

Should LLM Safety Be More Than Refusing Harmful Instructions? CipherBank: Exploring the Boundary of LLM Reasoning Capabilities through Cryptography Challenges

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.317124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.317124Z digest=sha256:5d4d443a9d313f391da51eb42e9f1bbe7494b36f1b04c4116354252ba151a553

Observation d3919f68-8a83-4a68-bfdd-33df43cb7dea · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.325604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.325604Z digest=sha256:a819dac45d32fa0bfc461997ab7e847e2c17dfcb6b6ed9f856239d354bf39fc3

Observation b091d066-6201-4946-a08c-2e603dcaceda · outbound

This paper cites CodeChameleon: Personalized Encryption Framework for Jailbreaking Large Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? CodeChameleon: Personalized Encryption Framework for Jailbreaking Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.338065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.338065Z digest=sha256:6eca77ba51992cb2e40bbe9265d50f616ca7aff4b1483ce6126d62be87a5f3a1

Observation 642d42cb-e567-4d53-81b7-f1669b2ff556 · outbound

This paper cites Benchmarking Large Language Models for Cryptanalysis and Side-Channel Vulnerabilities.

Should LLM Safety Be More Than Refusing Harmful Instructions? Benchmarking Large Language Models for Cryptanalysis and Side-Channel Vulnerabilities

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.346115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.346115Z digest=sha256:e1991f87a7db8ca80e35687fa4e4bef6882b2dff05c4365bb9927c77c8270da3

Observation 855c9b16-6bc5-494d-a596-913432d3e6df · outbound

This paper cites SaRO: Enhancing LLM Safety through Reasoning-based Alignment.

Should LLM Safety Be More Than Refusing Harmful Instructions? SaRO: Enhancing LLM Safety through Reasoning-based Alignment

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.354047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.354047Z digest=sha256:f16a98f0cfab6f98f28294b57541a171481d040c7ca7365daef037e5804608ca

Observation e83d5f35-8bbc-474d-88e1-43945ded8fd0 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 30

Resolution
verified exact
raw_fallback, observed 2026-08-07T11:27:48.173546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.369437Z digest=sha256:123ec0d19e5d7422a8f96712d7d1cb6568ff563ddfd7aad939ec1cf2b7c09fa5

Observation e64ce843-07d1-48f9-a367-3428c3f2529b · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.375342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.375342Z digest=sha256:5294b7530dfc13cfcab1964a62970fda71c2e7dd612dc0b40d6a979771572f44

Observation ccd256c9-e1ec-46b8-ada0-20f9e9743acd · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.381480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.381480Z digest=sha256:b4a9a67adbd8a8db432770691fcd81203c66e082a00210966f77fa68e990248d

Observation cf9f0039-c899-41fd-96e7-deba8fd39018 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.399243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.386841Z digest=sha256:bb5e0b8a966ceb749a2b60135bdac9bc2a71478e619eaa7929f85b4729522ffe

Observation 97a19b08-9b36-452a-b4f3-21651ea0bb56 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.393410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.393410Z digest=sha256:37ce577de3cbe42690229e4b6f209b7c550948f33212999c7c3f08263e87248f

Observation 9b9643af-eeed-44d3-97cf-291528fbb9bd · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.272487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.399564Z digest=sha256:794022b9711a3de0d60613db57d884f9fcebc16e7737804037bc858b9711dd65

Observation 28416321-807b-4036-978f-a1716094620e · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Should LLM Safety Be More Than Refusing Harmful Instructions? LLaMA: Open and Efficient Foundation Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.405477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.405477Z digest=sha256:d830bc66927f5fddcd74389d6f513a98a8090d8b3e5995f317e6ded1b7e59e04

Observation 3299d20c-9e50-4068-9f5b-e6512fbb7019 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.412052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.412052Z digest=sha256:c0040d6f42470bc1d65d1df2d761e1d1c807fc3ed7195e6c8b612ab301c7ce8d

Observation 83349e79-5cd6-4017-b2b7-2d15e981583f · outbound

This paper cites Dai, and Quoc V Le.

Should LLM Safety Be More Than Refusing Harmful Instructions? Dai, and Quoc V Le

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.418202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.418202Z digest=sha256:206a727f348206ff1d8458e8eb2ad9a007b70e992e80dd526c6f55322978fc3f

Observation 210a72bd-66f7-41e1-b106-a2140787b95f · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.424467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.424467Z digest=sha256:96800ab9f1b08e102357bbbf96cd3c397df704a5323beda489ded02fb860ee30

Observation cfd2488b-dc86-4db2-b031-d07166648d10 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:27:49.018405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T11:27:47.430319Z digest=sha256:41955d758dd805d3d4eb89047ba1b9cc966dd7278f7c22c1c2908b95d83b35da

Observation 3f819ca3-cc5c-413f-bf67-8f459f50fc71 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.435792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.435792Z digest=sha256:f29b4bd84211e59c9f8125806ff5e94eac530596cdc4c3c6e7d576828c04c06f

Observation 309629fa-7449-4a7c-9598-8fb71354e27a · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.443024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.443024Z digest=sha256:dac502fb4ad863a13171115e2243222120b5ac057dedd909831459068984e988

Observation b54cedb5-02a6-4160-b10b-f12f9acbbc3f · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.449114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.449114Z digest=sha256:b9eb04145e11931cc0c69bf2f83cfd75141bca727fd1745a47538afbe71843c3

Observation c32352f6-5eb5-42a8-8ed4-8f8deedc95e8 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.454913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.454913Z digest=sha256:a754c833d27c59db0c6f8b4097f700abd39da066433d118c08bdde0d9a10a0de

Observation 4e57b8c6-f9cd-4a4b-b2f0-07c74fd00440 · outbound

This paper cites an unresolved cited work.

Should LLM Safety Be More Than Refusing Harmful Instructions? Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.461431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.461431Z digest=sha256:4cab8e605fa770397ae01a96d33695021e75965d26737760250d474d24f1295e

Observation 86e2506d-b829-49ed-8cf4-cd50cc697115 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Should LLM Safety Be More Than Refusing Harmful Instructions? BERTScore: Evaluating Text Generation with BERT

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.467267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.467267Z digest=sha256:3e9cd9501eca411dc2e93953e5d63a087c5af2f9cdd7c225f97392620205d280

Observation 411437fe-d19a-48b6-a933-82d84bff3c92 · outbound

This paper cites online" 'onlinestring :=.

Should LLM Safety Be More Than Refusing Harmful Instructions? online" 'onlinestring :=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.475035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.475035Z digest=sha256:5d9c4f80519f10b777e8b208038cb726122ed2d258187efd81b8b360c5885af7

Observation 4d16e431-eaa6-413b-a14e-88b605b2ba9d · outbound

This paper cites write newline.

Should LLM Safety Be More Than Refusing Harmful Instructions? write newline

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:47.482807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:27:47.482807Z digest=sha256:d7817f25bec78b8d97fc32919c1a794466770ca4ef2795fa29be1bff3bcb5e51

Pith citing papers

No inbound Pith citation observations are available.