Pith. sign in

Paper Citation Record · LEDGER

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law

As of 8 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2506.06391.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06391 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:55.461862Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact10
  • verified fuzzy6
  • unresolved18
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 574a406c-f831-4253-9fb6-7e3a53640357 · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.334141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.334141Z digest=sha256:e6fba3bd2b0c9a036c7a70522ec8c56bc67b2ac38985a3081da26efe9565be49

Observation 60a6e142-cd78-4ee0-b820-07d5935e4212 · outbound

This paper cites The Radicalization Risks of GPT-3 and Advanced Neural Language Models.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law The Radicalization Risks of GPT-3 and Advanced Neural Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.406815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.406815Z digest=sha256:136a0439c83fd705c56bf9a7901e9f6d131bbb762ba0cc5524eabe36959648df

Observation 28f4630a-852f-4e1f-8119-cd5238a65cdf · outbound

This paper cites Klyman, ‘Acceptable Use Policies for Foundation Models’, Proceedings of the AAAI/ACM Conference on AI, Ethics, and So- ciety, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Klyman, ‘Acceptable Use Policies for Foundation Models’, Proceedings of the AAAI/ACM Conference on AI, Ethics, and So- ciety, vol

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.513958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.513958Z digest=sha256:1731ced5b5dc02c6f438a3a828ad21f9b9486f47435aa539dc6fbf5a747ba317

Observation 4358394e-7cca-4282-b6d5-daa2227aa4db · outbound

This paper cites Henckaerts and L.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Henckaerts and L

Reference 4

Resolution
verified exact
doi, observed 2026-08-07T10:28:57.751025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:52.598217Z digest=sha256:a8441aefd4b604b37a29ddef73daa7a5b1f85431fada19c0149c79d3154e987c

Observation 7ac05577-70b5-4328-ab54-c0e57ca6e660 · outbound

This paper cites As an AI language model, I cannot.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law As an AI language model, I cannot

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.675257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.675257Z digest=sha256:70b930599f0192b8b2329e1acb7e804bdb092a3e001ea6f09ab7060d33aae356

Observation 8c19d996-7c95-40ef-8887-1083d2e49e11 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Constitutional AI: Harmlessness from AI Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.769177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.769177Z digest=sha256:4b96909b0ba5e3fc65091bbfb6714206a406db8dbbb42ad9b3d723d46f5f964e

Observation 34896cab-3e45-46f0-b59a-f53c3fcb04d0 · outbound

This paper cites Don't Say No: Jailbreaking LLM by Suppressing Refusal.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.855847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.855847Z digest=sha256:71d613d597e3f83f2432992295d31005ac79a2f8b644e7f5bd0e54c15f82ba00

Observation e2ce047c-3d2a-4395-bd65-600f955b599d · outbound

This paper cites Milaninia, ‘Biases in machine learning models and big data analytics: The inter- national criminal and humanitarian law im- plications’, Int.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Milaninia, ‘Biases in machine learning models and big data analytics: The inter- national criminal and humanitarian law im- plications’, Int

Reference 8

Resolution
verified exact
doi, observed 2026-08-07T10:28:57.481074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:52.991890Z digest=sha256:43c958b9cc1c483de8e37d748fdd2f1e249515e3eabb407795da64ce66b0b428

Observation 8443e5a1-ed35-40cf-8bee-b3ceebed01cd · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:29:00.820672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.072596Z digest=sha256:bf4e2e48423061a54bc9ef2cff8059a004099c861856809f451dd10eca3b826f

Observation 45affa65-05f6-4f1c-bf7a-c1414ae12f0b · outbound

This paper cites Marcos, ‘Can large language models apply the law?’, AI and Society, pp.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Marcos, ‘Can large language models apply the law?’, AI and Society, pp

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T10:28:57.323876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.208628Z digest=sha256:e22b0cee73edfad897286bc7922f615b28c7ddf15ffe872853a5f149667213f4

Observation 12450b47-255a-49f7-ac2c-8d758ae66a99 · outbound

This paper cites Available: https://www.ohchr.org/en/instruments-and- mechanisms/international-human-rights-law.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Available: https://www.ohchr.org/en/instruments-and- mechanisms/international-human-rights-law

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:00.597138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.257276Z digest=sha256:ca4e7bc6d2f277d856907eaefff73d800ad2a49a61c7d922d0711ddc67826861

Observation 3e759a0d-f949-4320-bb48-db40320f9f66 · outbound

This paper cites Available: https://www.ohchr.org/en/resources/educato rs/human-rights-education-training/universal- declaration-human-rights-1948.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Available: https://www.ohchr.org/en/resources/educato rs/human-rights-education-training/universal- declaration-human-rights-1948

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:00.270595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.330652Z digest=sha256:74f6d6d4fb6b86ba2836ec0f2efe4c49ff3b40f2a5babaacd96c8e70875cb0cf

Observation 3165029b-0832-4650-80b5-55033597d551 · outbound

This paper cites Available: https://civil-protection-humanitarian- aid.ec.europa.eu/what/humanitarian- aid/international-humanitarian-law.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Available: https://civil-protection-humanitarian- aid.ec.europa.eu/what/humanitarian- aid/international-humanitarian-law

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:00.037168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.410472Z digest=sha256:e677d0cb3f3f7ffc91d45dad87b3c5621f90513ff918c4d8106508f58b6b0300

Observation cd085da7-8493-4990-bbf8-e2f95ebbc257 · outbound

This paper cites MacLaren and F.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law MacLaren and F

Reference 14

Resolution
verified exact
doi, observed 2026-08-07T10:28:57.065163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.495867Z digest=sha256:88cf0e220622f22f62f1b574ef48c5bc695e67114a0ca60425f7947aa83679fd

Observation 5d3d0bd1-5305-4e1f-a526-c2f80c6b4fe7 · outbound

This paper cites Zhang, M.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Zhang, M

Reference 15

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.915104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.548232Z digest=sha256:90b69bb74360198aad2b7a7ec486c481b251fd65504673f26a530cae24cd850c

Observation 65ccd62e-d953-4566-8f4a-68a419aa1c14 · outbound

This paper cites Refusal Behavior in Large Language Models: A Nonlinear Perspective.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Refusal Behavior in Large Language Models: A Nonlinear Perspective

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:53.629021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:53.629021Z digest=sha256:c4ff69a47afe1c654bb6bee172df228c53c57adbf7a841050c7254cf6773f018

Observation f9e61acd-30ff-4dbb-9931-d312de3603c9 · outbound

This paper cites Does Refusal Training in LLMs Generalize to the Past Tense?.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Does Refusal Training in LLMs Generalize to the Past Tense?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:53.683286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:53.683286Z digest=sha256:0afc9692e01c560d05ab25dde7efb14206fee5512c0c398240511ca5291add97

Observation 4ca58ba4-11fb-4382-a51a-9873aae6d3e4 · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:59.670807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.775394Z digest=sha256:f6b1b9f8efb4bbd2935ea91befc88d579a3fe9813722f9b7dab214e1afc1185e

Observation ddeb866e-b736-45a1-ba20-8621612a8231 · outbound

This paper cites Pomson, ‘Methodology of identifying cus- tomary international law applicable to cyber activities’, LeidenJournalofInternationalLaw, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Pomson, ‘Methodology of identifying cus- tomary international law applicable to cyber activities’, LeidenJournalofInternationalLaw, vol

Reference 19

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.616471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.880018Z digest=sha256:9344cd52a763a05188033359b4a0cb9b155159f047ebb41fb9384396b095f60d

Observation 40d66404-5b12-4fde-ba51-478a18715cb4 · outbound

This paper cites Yudkowsky, ‘The AI Alignment Prob- lem: Why It’s Hard, and Where to Start’.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Yudkowsky, ‘The AI Alignment Prob- lem: Why It’s Hard, and Where to Start’

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:59.339063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:53.958238Z digest=sha256:48b9564cfd161531c3020ecd7c8d2a18998bb2cd3c0fd2c6b088ce382f58fbfc

Observation 315cca8b-bb6b-4a48-8416-01e8c44dad0e · outbound

This paper cites Gabriel, ‘Artificial Intelligence, Values, and Alignment’, Minds and Machines, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Gabriel, ‘Artificial Intelligence, Values, and Alignment’, Minds and Machines, vol

Reference 21

Resolution
malformed identifier
no resolver link, observed 2026-08-07T10:28:54.089537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.089537Z digest=sha256:ee8b244548b9f344f19abbf7b7983053f6ccc226b86581209b0eda62018cd69f

Observation 25992e24-38b4-4f99-bd6d-29bf0b781559 · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:58.947214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:54.214140Z digest=sha256:f881d75c44ce1e04839e78ac548dc476c84ff226f52626efbc6198ebcbb84691

Observation 20671a0a-4bb0-4bc0-bdfb-d614d4463660 · outbound

This paper cites an unresolved cited work.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:58.661986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:54.289250Z digest=sha256:fa56a58baaa7a0235991ce99fd9c62c80705bf680efb92eee2f6fb2492af95c0

Observation aeea0786-0e18-48f6-9888-f708798daa3b · outbound

This paper cites International Scientific Report on the Safety of Advanced AI (Interim Report).

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law International Scientific Report on the Safety of Advanced AI (Interim Report)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:54.373523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.373523Z digest=sha256:17a5ede6cda5871aa695922ec2aaeaf648e2e9b081c59dcd151fe30cf86c9f22

Observation e4130bbd-beee-4440-986e-657d2eec1104 · outbound

This paper cites Strzępek, ‘Human Rights as a Factor in the AI Alignment’, GIS Odyssey Journal, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Strzępek, ‘Human Rights as a Factor in the AI Alignment’, GIS Odyssey Journal, vol

Reference 25

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.407578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:54.475083Z digest=sha256:28db5629c7d89dfdb2a8e10508841f30469e66c098ea60d7611804ae19837dc9

Observation 4f2251a5-a415-4c12-b106-67f98b8ce38b · outbound

This paper cites Szpor, ‘European Legal Framework for the Use of Artificial Intelligence in Pub- licly Accessible Space’, GIS Odyssey Journal, vol.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Szpor, ‘European Legal Framework for the Use of Artificial Intelligence in Pub- licly Accessible Space’, GIS Odyssey Journal, vol

Reference 26

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.199379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:54.569206Z digest=sha256:33eacd7f5143313abb8b6a2782c686fc8df629b29ab6a92ba1f41a3661a792b4

Observation a0548506-e31e-4e21-91d2-9ec7121100b2 · outbound

This paper cites Novelli, F.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Novelli, F

Reference 27

Resolution
malformed identifier
no resolver link, observed 2026-08-07T10:28:54.645387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.645387Z digest=sha256:52021f51b7466c55fbed785b3fd4db7fff9c862820b415611a8edcb7690a1e97

Observation eef5aa8c-12ea-43a2-b88e-362704391942 · outbound

This paper cites Ji et al., ‘BeaverTails: Towards Im- proved Safety Alignment of LLM via a Human-Preference Dataset’, Nov.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Ji et al., ‘BeaverTails: Towards Im- proved Safety Alignment of LLM via a Human-Preference Dataset’, Nov

Reference 28

Resolution
verified exact
doi, observed 2026-08-07T10:28:56.024830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:54.719928Z digest=sha256:37b3d8590a6d77fe888c1f86cc70ab0f49179753d3c0e0e0ff489367b3a64347

Observation c0f879a3-92f5-496f-abec-9b8f05638b4d · outbound

This paper cites ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:54.802416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.802416Z digest=sha256:48e3342fbb74225da5d9b68a51723713e66a9e7bdfaebdda17d90d3b6528ea7c

Observation 5e333f84-534b-47cf-826d-86effaff999f · outbound

This paper cites https://platform.openai.com.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law https://platform.openai.com

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:58.482180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:54.893392Z digest=sha256:a64c417cb6b28799aec49126f9387d55d4ed163dfc78b8bcdf7382dca6c2835a

Observation 14de37fd-3c28-4821-adbe-efc2d4d0a818 · outbound

This paper cites OR-Bench: An Over-Refusal Benchmark for Large Language Models.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law OR-Bench: An Over-Refusal Benchmark for Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:54.986950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:54.986950Z digest=sha256:5d9ca12b53063086dc33cd6fcc1617db259f0ec478d55849347b3dc233745a38

Observation ea34ee52-42ec-4c1e-ae5d-96a5e1f37298 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:55.100445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:55.100445Z digest=sha256:cfacdb182c30d7a9d8c89d17726bc42b98fcdca4d41998ea70d2e972597c5f29

Observation 19f8fb8a-0a31-4957-a2da-2040eaa170a2 · outbound

This paper cites ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:55.208535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:55.208535Z digest=sha256:7b8210f21fe54f6d425993e5a14954de97936c102a4b10361a167c4fad55319f

Observation af9d2d6b-db8e-4234-86da-380e59faee2c · outbound

This paper cites Stanovsky, R.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Stanovsky, R

Reference 34

Resolution
verified exact
doi, observed 2026-08-07T10:28:55.714860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:55.290535Z digest=sha256:847086fec214bf321db8bc3b6a6f6fd35820cc17d141beb57130093e2069a25a

Observation c6970e34-3663-4ae1-9c1e-4db664fc4553 · outbound

This paper cites Scheutz, R.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Scheutz, R

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:55.374050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:55.374050Z digest=sha256:23082e2ad6da180731889fa7f669e4c24ded8c696569ef3c06ed81057389067d

Observation 9ed424fb-6b34-4866-ac7e-41c3de476ada · outbound

This paper cites Claude 3.7 system card.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Claude 3.7 system card

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:58.271519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:28:55.461862Z digest=sha256:faa6ee6ce6ffdacc3e163825fe04cce913c97eed63538fca564dbfb50cc6454f

Pith citing papers

No inbound Pith citation observations are available.