Pith. sign in

Paper Citation Record · LEDGER

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law

As of 9 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2506.06391.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06391 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:55.461862Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact10
  • verified fuzzy6
  • unresolved18
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 574a406c-f831-4253-9fb6-7e3a53640357 · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.334141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.334141Z digest=sha256:e6fba3bd2b0c9a036c7a70522ec8c56bc67b2ac38985a3081da26efe9565be49

Observation 60a6e142-cd78-4ee0-b820-07d5935e4212 · outbound

This paper cites The Radicalization Risks of GPT-3 and Advanced Neural Language Models.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law The Radicalization Risks of GPT-3 and Advanced Neural Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.406815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.406815Z digest=sha256:136a0439c83fd705c56bf9a7901e9f6d131bbb762ba0cc5524eabe36959648df

Observation 28f4630a-852f-4e1f-8119-cd5238a65cdf · outbound

This paper cites Klyman, ‘Acceptable Use Policies for Foundation Models’, Proceedings of the AAAI/ACM Conference on AI, Ethics, and So- ciety, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Klyman, ‘Acceptable Use Policies for Foundation Models’, Proceedings of the AAAI/ACM Conference on AI, Ethics, and So- ciety, vol

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.513958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.513958Z digest=sha256:1731ced5b5dc02c6f438a3a828ad21f9b9486f47435aa539dc6fbf5a747ba317

Observation 4358394e-7cca-4282-b6d5-daa2227aa4db · outbound

This paper cites Henckaerts and L.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Henckaerts and L

Reference 4

Resolution
verified exact
doi, observed 2026-08-07T10:28:57.751025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:52.598217Z digest=sha256:ff4a3f96e9d2015d0709d89544d2df327ca16619cdccbc229fd15c364b940a72

Observation 7ac05577-70b5-4328-ab54-c0e57ca6e660 · outbound

This paper cites As an AI language model, I cannot.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law As an AI language model, I cannot

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.675257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.675257Z digest=sha256:70b930599f0192b8b2329e1acb7e804bdb092a3e001ea6f09ab7060d33aae356

Observation 8c19d996-7c95-40ef-8887-1083d2e49e11 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Constitutional AI: Harmlessness from AI Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.769177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.769177Z digest=sha256:4b96909b0ba5e3fc65091bbfb6714206a406db8dbbb42ad9b3d723d46f5f964e

Observation 34896cab-3e45-46f0-b59a-f53c3fcb04d0 · outbound

This paper cites Don't Say No: Jailbreaking LLM by Suppressing Refusal.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.855847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.855847Z digest=sha256:71d613d597e3f83f2432992295d31005ac79a2f8b644e7f5bd0e54c15f82ba00

Observation e2ce047c-3d2a-4395-bd65-600f955b599d · outbound

This paper cites Milaninia, ‘Biases in machine learning models and big data analytics: The inter- national criminal and humanitarian law im- plications’, Int.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Milaninia, ‘Biases in machine learning models and big data analytics: The inter- national criminal and humanitarian law im- plications’, Int

Reference 8

Resolution
verified exact
doi, observed 2026-08-07T10:28:57.481074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:52.991890Z digest=sha256:a9dca9e8f5f2346fe031d0ad9d076460a320de207b043729f87b214ac2889893

Observation 8443e5a1-ed35-40cf-8bee-b3ceebed01cd · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:29:00.820672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.072596Z digest=sha256:bd9570b8f6d521f1cfb1482d3957f5d36f1e0351e9a8d34f020ea8754e9fd5df

Observation 45affa65-05f6-4f1c-bf7a-c1414ae12f0b · outbound

This paper cites Marcos, ‘Can large language models apply the law?’, AI and Society, pp.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Marcos, ‘Can large language models apply the law?’, AI and Society, pp

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T10:28:57.323876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.208628Z digest=sha256:c7b49770f294270f6e09983b4437fce5c70a723fbfd0a23af0eea6a294ca53f8

Observation 12450b47-255a-49f7-ac2c-8d758ae66a99 · outbound

This paper cites Available: https://www.ohchr.org/en/instruments-and- mechanisms/international-human-rights-law.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Available: https://www.ohchr.org/en/instruments-and- mechanisms/international-human-rights-law

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:00.597138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.257276Z digest=sha256:76e54878c1f6b20ed445cac348a256d76783f13eca9520306235bc4b02d4a41b

Observation 3e759a0d-f949-4320-bb48-db40320f9f66 · outbound

This paper cites Available: https://www.ohchr.org/en/resources/educato rs/human-rights-education-training/universal- declaration-human-rights-1948.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Available: https://www.ohchr.org/en/resources/educato rs/human-rights-education-training/universal- declaration-human-rights-1948

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:00.270595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.330652Z digest=sha256:78d7b6f22024c0b8a63771bb821e012d288ecab09c7df37dc589fbc2df2c506e

Observation 3165029b-0832-4650-80b5-55033597d551 · outbound

This paper cites Available: https://civil-protection-humanitarian- aid.ec.europa.eu/what/humanitarian- aid/international-humanitarian-law.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Available: https://civil-protection-humanitarian- aid.ec.europa.eu/what/humanitarian- aid/international-humanitarian-law

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:00.037168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.410472Z digest=sha256:a8aae789edeeb09fbaa00f45b6ac9b30d6627a97d3cb56f24a12c9878e211fc2

Observation cd085da7-8493-4990-bbf8-e2f95ebbc257 · outbound

This paper cites MacLaren and F.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law MacLaren and F

Reference 14

Resolution
verified exact
doi, observed 2026-08-07T10:28:57.065163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.495867Z digest=sha256:783944f9bd9461d279497c2b03c6c4f6bcec7103573e2cb5f3ea75ae6c82adcb

Observation 5d3d0bd1-5305-4e1f-a526-c2f80c6b4fe7 · outbound

This paper cites Zhang, M.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Zhang, M

Reference 15

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.915104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.548232Z digest=sha256:121b25950757dfc3d6b887f49eff21ca51576bf7e2259201c3790392ccebfe86

Observation 65ccd62e-d953-4566-8f4a-68a419aa1c14 · outbound

This paper cites Refusal Behavior in Large Language Models: A Nonlinear Perspective.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Refusal Behavior in Large Language Models: A Nonlinear Perspective

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:53.629021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:53.629021Z digest=sha256:c4ff69a47afe1c654bb6bee172df228c53c57adbf7a841050c7254cf6773f018

Observation f9e61acd-30ff-4dbb-9931-d312de3603c9 · outbound

This paper cites Does Refusal Training in LLMs Generalize to the Past Tense?.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Does Refusal Training in LLMs Generalize to the Past Tense?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:53.683286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:53.683286Z digest=sha256:0afc9692e01c560d05ab25dde7efb14206fee5512c0c398240511ca5291add97

Observation 4ca58ba4-11fb-4382-a51a-9873aae6d3e4 · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:59.670807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.775394Z digest=sha256:2ee16928ca24a44897ffd84a2bb5efc26301d502d9a4643ba8cc4d8aba84d410

Observation ddeb866e-b736-45a1-ba20-8621612a8231 · outbound

This paper cites Pomson, ‘Methodology of identifying cus- tomary international law applicable to cyber activities’, LeidenJournalofInternationalLaw, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Pomson, ‘Methodology of identifying cus- tomary international law applicable to cyber activities’, LeidenJournalofInternationalLaw, vol

Reference 19

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.616471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.880018Z digest=sha256:da15ed8cc411e8f8712f10a4ac4bd0340dae485b3c25a043d351816588090798

Observation 40d66404-5b12-4fde-ba51-478a18715cb4 · outbound

This paper cites Yudkowsky, ‘The AI Alignment Prob- lem: Why It’s Hard, and Where to Start’.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Yudkowsky, ‘The AI Alignment Prob- lem: Why It’s Hard, and Where to Start’

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:59.339063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:53.958238Z digest=sha256:2c465949c053f003a1a67332c1eaf62730b448611162f730705d812fbef01606

Observation 315cca8b-bb6b-4a48-8416-01e8c44dad0e · outbound

This paper cites Gabriel, ‘Artificial Intelligence, Values, and Alignment’, Minds and Machines, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Gabriel, ‘Artificial Intelligence, Values, and Alignment’, Minds and Machines, vol

Reference 21

Resolution
malformed identifier
no resolver link, observed 2026-08-07T10:28:54.089537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.089537Z digest=sha256:ee8b244548b9f344f19abbf7b7983053f6ccc226b86581209b0eda62018cd69f

Observation 25992e24-38b4-4f99-bd6d-29bf0b781559 · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:58.947214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:54.214140Z digest=sha256:499fa838299e44721bf86c529504a6793efdda2d49284400ad73e79be8afa7fd

Observation 20671a0a-4bb0-4bc0-bdfb-d614d4463660 · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:58.661986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:54.289250Z digest=sha256:06fc73a23eb1b7ea025f6c2c71bda8dbaa871f65990dc8f001414a64a1f9ea21

Observation aeea0786-0e18-48f6-9888-f708798daa3b · outbound

This paper cites International Scientific Report on the Safety of Advanced AI (Interim Report).

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law International Scientific Report on the Safety of Advanced AI (Interim Report)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:54.373523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.373523Z digest=sha256:17a5ede6cda5871aa695922ec2aaeaf648e2e9b081c59dcd151fe30cf86c9f22

Observation e4130bbd-beee-4440-986e-657d2eec1104 · outbound

This paper cites Strzępek, ‘Human Rights as a Factor in the AI Alignment’, GIS Odyssey Journal, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Strzępek, ‘Human Rights as a Factor in the AI Alignment’, GIS Odyssey Journal, vol

Reference 25

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.407578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:54.475083Z digest=sha256:a912c5c1d58d54bcba56aec2a34d896c8018c7838bd23601b39d5c9fd1fb4271

Observation 4f2251a5-a415-4c12-b106-67f98b8ce38b · outbound

This paper cites Szpor, ‘European Legal Framework for the Use of Artificial Intelligence in Pub- licly Accessible Space’, GIS Odyssey Journal, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Szpor, ‘European Legal Framework for the Use of Artificial Intelligence in Pub- licly Accessible Space’, GIS Odyssey Journal, vol

Reference 26

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.199379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:54.569206Z digest=sha256:a4878cd116088717d9889f4a4d1fa8fd9824cc1693cea3b9bdec2a600f41a18f

Observation a0548506-e31e-4e21-91d2-9ec7121100b2 · outbound

This paper cites Novelli, F.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Novelli, F

Reference 27

Resolution
malformed identifier
no resolver link, observed 2026-08-07T10:28:54.645387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.645387Z digest=sha256:52021f51b7466c55fbed785b3fd4db7fff9c862820b415611a8edcb7690a1e97

Observation eef5aa8c-12ea-43a2-b88e-362704391942 · outbound

This paper cites Ji et al., ‘BeaverTails: Towards Im- proved Safety Alignment of LLM via a Human-Preference Dataset’, Nov.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Ji et al., ‘BeaverTails: Towards Im- proved Safety Alignment of LLM via a Human-Preference Dataset’, Nov

Reference 28

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.024830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:54.719928Z digest=sha256:0f735cf8e395ab64e4d3c2177f0cefaa42c800853023afbf95bfb430abadbcd4

Observation c0f879a3-92f5-496f-abec-9b8f05638b4d · outbound

This paper cites ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:54.802416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.802416Z digest=sha256:48e3342fbb74225da5d9b68a51723713e66a9e7bdfaebdda17d90d3b6528ea7c

Observation 5e333f84-534b-47cf-826d-86effaff999f · outbound

This paper cites https://platform.openai.com.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law https://platform.openai.com

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:58.482180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:54.893392Z digest=sha256:641a2793d5cbe4706afdc1e8a625c0826dc78e1449cb80afd1ba3d520866de8c

Observation 14de37fd-3c28-4821-adbe-efc2d4d0a818 · outbound

This paper cites OR-Bench: An Over-Refusal Benchmark for Large Language Models.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law OR-Bench: An Over-Refusal Benchmark for Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:54.986950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.986950Z digest=sha256:5d9ca12b53063086dc33cd6fcc1617db259f0ec478d55849347b3dc233745a38

Observation ea34ee52-42ec-4c1e-ae5d-96a5e1f37298 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:55.100445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:55.100445Z digest=sha256:cfacdb182c30d7a9d8c89d17726bc42b98fcdca4d41998ea70d2e972597c5f29

Observation 19f8fb8a-0a31-4957-a2da-2040eaa170a2 · outbound

This paper cites ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:55.208535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:55.208535Z digest=sha256:7b8210f21fe54f6d425993e5a14954de97936c102a4b10361a167c4fad55319f

Observation af9d2d6b-db8e-4234-86da-380e59faee2c · outbound

This paper cites Stanovsky, R.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Stanovsky, R

Reference 34

Resolution
verified exact
doi, observed 2026-08-07T10:28:55.714860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:55.290535Z digest=sha256:7a2561312df5e38fde290270b627ca972654a0dd307ac8de5fada0a780c083ff

Observation c6970e34-3663-4ae1-9c1e-4db664fc4553 · outbound

This paper cites Scheutz, R.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Scheutz, R

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:55.374050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:55.374050Z digest=sha256:23082e2ad6da180731889fa7f669e4c24ded8c696569ef3c06ed81057389067d

Observation 9ed424fb-6b34-4866-ac7e-41c3de476ada · outbound

This paper cites Claude 3.7 system card.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Claude 3.7 system card

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:58.271519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:28:55.461862Z digest=sha256:275df612bf5501972da07754d2343580965da292595d822cbd3e01a52cf55ad1

Pith citing papers

No inbound Pith citation observations are available.