Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:34:36.614045Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:2501.08246.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:34:36.614045Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-26T04:35:51.583460Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
45 of 45 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 54713b1b-b6a0-4c6c-b60f-ea6166841c16 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4e231c4-3d29-4810-a314-3e3ad938df36 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f66ff6c-523e-4d0b-afda-50202acd4d96 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints GPT-4 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b63ca7f-11d7-4f89-b660-ec083cdd001e · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 90329bb6-e5bc-40ea-b4b3-1cbf8b197884 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints D.; Ho, J.; Tarlow, D.; and Van Den Berg, R
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ea9cedbe-ee2a-4934-97d3-d2d2149f723d · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Constitutional AI: Harmlessness from AI Feedback
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df5430e5-3588-46bd-861c-d3a40ab04bbc · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Training Diffusion Models with Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5340d82a-18b8-4cef-977d-3eee2b8506de · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 365c9c4b-56d9-4845-8b57-59c53a93cc69 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1ba8fc20-1883-4e38-8485-745a623cc97a · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a0dacecf-ce7b-4d63-a703-1670304097b0 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f0d84e8-bcad-4345-a5de-bab64235c621 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a508ff64-cec2-47e7-8263-495216ba1dcf · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42437c03-9628-4384-9752-c57b2421d182 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints R.; Srivastava, A.; and Agrawal, P
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1cee21ef-7d17-4b59-ad84-6728e0ee1c14 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Baseline Defenses for Adversarial Attacks Against Aligned Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8fd3733-4d92-461d-9447-6f7fc399d0d7 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b7c7188b-c6f3-4c8e-9eaa-4aa8181e4aaf · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ba88c8c3-1e8b-4b93-8a26-bc931f53e8bf · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Certifying LLM Safety against Adversarial Prompting
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54d2cb09-86ff-4275-bceb-8c9463244fa4 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Open Sesame! Universal Black Box Jailbreaking of Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcca81fd-e948-4fc2-824c-4aaca73527ce · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints RAIN: Your Language Models Can Align Themselves without Finetuning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59dac234-869a-4c30-84a2-d8250452abd2 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 97868760-a7c4-4081-b7bf-50b7c8e79689 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5912a16d-6e82-4c98-bdcd-d2ffc1d317c0 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2fe14311-d545-4959-9c7f-6509a9fe045c · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints FLIRT: Feedback Loop In-context Red Teaming
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f5f3076-d12d-4be7-877d-a3cb721e7ee2 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Text Embeddings Reveal (Almost) As Much As Text
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b6d4d1a-eb49-4399-9321-9c0f22a9660a · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2bad4ddf-be31-4a38-83f2-4a40210dcb9f · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Instruction Tuning with GPT-4
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1292782a-a249-42d1-8d14-741e0fac3c20 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Red Teaming Language Models with Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39040c31-ea35-4e10-bf47-ccc48889ea73 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints D.; Ermon, S.; and Finn, C
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0b5e7d15-7e94-4a22-9f8b-cf62af5066d1 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc77d5d7-d66b-416c-9a9d-16412cad66f4 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb8bb8e6-3b15-4f39-b420-5965c6196b17 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints High-Resolution Image Synthesis with Latent Diffusion Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 161187d3-57f3-4263-b381-71953cf35b88 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Code Llama: Open Foundation Models for Code
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22babea3-4a10-486b-9d02-f0b1a4cd12dd · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Proximal Policy Optimization Algorithms
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92c0e8a5-cfd6-4441-83d2-11bc8389e9a8 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints CodeFusion: A Pre-trained Diffusion Model for Code Generation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 596bda71-a020-47ce-a070-baf05566a54d · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9c6e9b4e-ab85-4bb0-b9e7-408b60693cfe · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9f0d3bbc-1be3-408c-8337-e4e944dcfdb1 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 408691c8-4fcf-47c4-992b-30b4d11bb299 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 32e7b1e0-9d08-46f3-aaed-e34e0bcf68c3 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation de2478ee-3330-46b2-a4d4-58b0c3350eda · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Gradient-Based Language Model Red Teaming
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b95ecf7d-e00e-4012-a4b9-cbdb8aefe43f · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Low-Resource Languages Jailbreak GPT-4
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e97bd6b-7191-4497-8b0e-d60ff53edae0 · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a1cd0671-ac73-47b2-9da3-611800a9012a · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5679ae28-7e14-4d6e-9ab1-7a0b64e619dd · outbound
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b94eda2-1069-43b9-baa8-93564a4f492f · inbound
Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.