Pith. sign in

Paper Citation Record · LEDGER

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense

As of 13 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2608.10933.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10933 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:01:41.590694Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3621170c-3fdd-46e2-80f0-fc40c4eed616 · outbound

This paper cites luma dream machine, 2024.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense luma dream machine, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.862468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.251423Z digest=sha256:351880aa3898a6eeda848dc8c77256fb9b859a6cad4ea01d016a278b118db0b2

Observation ec10b06b-8c33-4795-9427-3d480ba6238c · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval, 2021.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Frozen in time: A joint video and image encoder for end-to-end retrieval, 2021

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.841263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.258136Z digest=sha256:fe729ea9854813599b024ef90524d835b6512e38de69e987da0d78deb6693108

Observation 3b67e2fa-adaa-4280-968c-b8d964b66e78 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.263970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.263970Z digest=sha256:3d6e7e8ed8a0aaa4127f301178d55a9d5bb92ceeb74b7c1a17ad66381f354d8d

Observation f00a8dc0-52d0-42c0-b2e2-9dc01d3ef0d5 · outbound

This paper cites Transformer Interpretability Beyond Attention Visualization.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Transformer Interpretability Beyond Attention Visualization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.272247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.272247Z digest=sha256:0580ab1f120ffc6bd33edc90ddc385192d47ed0482c19c7a708125e667c014dd

Observation 3207e168-661f-4923-8765-b9c2cd0cbc9d · outbound

This paper cites Transformer inter- pretability beyond attention visualization.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Transformer inter- pretability beyond attention visualization

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.820844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.278872Z digest=sha256:ccecf48a9283f638a7634647424b8afa3c9000411a279940ad7059b000b47d98

Observation dd18a8dc-8fa7-41f5-b1ef-592bd7b2418c · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffu- sion models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Videocrafter2: Overcoming data limitations for high-quality video diffu- sion models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.793770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.286281Z digest=sha256:71059e37aa9ef962c716cf95184ad0f75df9498cb09955915e993ab4ca34a0b7

Observation 1b194c8a-90e4-44d1-8fbb-0dd988a485f4 · outbound

This paper cites SafeWatch: An Efficient Safety-Policy Following Video Guardrail Model with Transparent Explanations.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense SafeWatch: An Efficient Safety-Policy Following Video Guardrail Model with Transparent Explanations

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.294275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.294275Z digest=sha256:2336e0ca69537e3fcc49b6cc632b4ff116871c2e5f01746ca8a258315b46c18c

Observation 6d9cfe28-841e-4496-95dc-fa4bf40d3c83 · outbound

This paper cites Fuzz-Testing Meets LLM-Based Agents: An Automated and Efficient Framework for Jailbreaking Text-To-Image Generation Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Fuzz-Testing Meets LLM-Based Agents: An Automated and Efficient Framework for Jailbreaking Text-To-Image Generation Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.299903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.299903Z digest=sha256:591091a878fa4d45927e1fc0c5bea288b47c17658a4d8550d380f805fd7f6634

Observation 3fd487d9-3984-4e6c-ac3f-90d1a8f2657f · outbound

This paper cites Not what you’ve signed up for: Compromising real-world llm-integrated ap- plications with indirect prompt injection.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Not what you’ve signed up for: Compromising real-world llm-integrated ap- plications with indirect prompt injection

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.776223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.307066Z digest=sha256:56b271acb829a2ac5af87cabc88bda99df47ac2b5b337699c680b89faee9b063

Observation c215ba6d-439a-4854-91f7-b8745555ad68 · outbound

This paper cites Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.313665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.313665Z digest=sha256:d8b9e90b6e9e5a6cb1b8ad70a13ca0c3b7c50ef02b036d4c2abe47ad4c6f7d66

Observation 66a3a2b8-8ce7-42e6-a674-8a2e9f4cc013 · outbound

This paper cites NoVo: Norm Voting off Hallucinations with Attention Heads in Large Language Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense NoVo: Norm Voting off Hallucinations with Attention Heads in Large Language Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-12T14:01:42.378552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.323593Z digest=sha256:cc53e3fcb0b8407f9ea905f832e794ecd29c3319123c5cdf028ea25403102e75

Observation 4c4bc3a6-4452-4519-b453-f790865937f6 · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.329399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.329399Z digest=sha256:f6782c65d7e7e956e8a4790d79f3d6f8e084d16c4017658c6a93d583d9fa982a

Observation 8972b9cb-ce23-4fc2-97e7-a8a31332e2be · outbound

This paper cites CogMorph: Cognitive Morphing Attacks for Text-to-Image Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense CogMorph: Cognitive Morphing Attacks for Text-to-Image Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.335157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.335157Z digest=sha256:cd68e8e280d41a8cdb4714e2e51cda30c1507fc2d68b3b816aaaa144555b33eb

Observation 13b9dec2-e8b1-4543-a0d5-68eeb5a64972 · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.341437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.341437Z digest=sha256:b99b4218e1deb42a6f15130a37ec101956aedbe84bd82b9f247e9d4995a8c596

Observation 43710e97-8d7a-4c1e-a991-9ea39212dba6 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.348391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.348391Z digest=sha256:407562ab5100fdafd84b9d22a93933570d690a7f2acaa58ea75e16b72e07c036

Observation 2d353691-0f5f-4120-91bc-cef5f852c683 · outbound

This paper cites Bridg- ing text and video generation: A survey.arXiv preprint arXiv:2510.04999, 2025.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Bridg- ing text and video generation: A survey.arXiv preprint arXiv:2510.04999, 2025

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.354134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.354134Z digest=sha256:7c2c58f17abb61ceb5cc87d0a02b443bb9066ce2da665c583f0eb8f6b9763e68

Observation bd25bb91-9d8f-4890-b3c5-17aa88224f50 · outbound

This paper cites Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.360266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.360266Z digest=sha256:7d7734f416ff1ada802c6568839fb4327aa15477fd036df8705488b5d09b61b9

Observation f784ee13-bc4b-41fc-a28e-37dd8c9be12e · outbound

This paper cites SafeLLM: Unlearning Harmful Outputs from Large Language Models against Jailbreak Attacks.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense SafeLLM: Unlearning Harmful Outputs from Large Language Models against Jailbreak Attacks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.366538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.366538Z digest=sha256:fa6e0979632c2cd1b6f5cd120e209d03d73a29740f267c11d2e59a9f36e89f0d

Observation d734b093-3c1e-468d-b241-58ca493f5254 · outbound

This paper cites Exploring inconsistent knowledge distil- lation for object detection with data augmentation.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Exploring inconsistent knowledge distil- lation for object detection with data augmentation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.753162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.371822Z digest=sha256:c42e1640487b987a5ac0aa0b95be033131ba0c5d6d4ab51044cc7450c3e82b49

Observation acb65aa9-fd16-4785-9da7-5838883f6ecd · outbound

This paper cites A large-scale multiple- objective method for black-box attack against object detec- tion.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense A large-scale multiple- objective method for black-box attack against object detec- tion

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.734396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.376752Z digest=sha256:f85e0bda45b7ecba9bf12c2792b8c585007ac48ea59db3732a319983c1d78596

Observation aa3512d5-d9fd-465e-a773-0bed35cd8693 · outbound

This paper cites Imitated detectors: Stealing knowl- edge of black-box object detectors.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Imitated detectors: Stealing knowl- edge of black-box object detectors

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.716384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.385652Z digest=sha256:1b57c32e01bd050cb41bf68be6829002e6ec499ed080b229de0799cf793199ec

Observation 5997afa1-715b-4658-93ed-45fe0e387139 · outbound

This paper cites Badclip: Dual- embedding guided backdoor attack on multimodal con- trastive learning.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Badclip: Dual- embedding guided backdoor attack on multimodal con- trastive learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.694850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.392269Z digest=sha256:b631136a7e892316bdc29cf15ac54b1f3371e09229acfae213ec63b88c06bd0c

Observation 03be24ac-c4ba-4923-95eb-0e8b77dbb5cb · outbound

This paper cites Re- visiting backdoor attacks against large vision-language mod- els from domain shift.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Re- visiting backdoor attacks against large vision-language mod- els from domain shift

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.675758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.403136Z digest=sha256:7c47577d1f9412fff32aea0746d0b7e39dfb3121cf3676311ba6a398222734ff

Observation 75b6184d-628f-4045-8014-aa06a4e4f31a · outbound

This paper cites T2VShield: Model-Agnostic Jailbreak Defense for Text-to-Video Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense T2VShield: Model-Agnostic Jailbreak Defense for Text-to-Video Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.409683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.409683Z digest=sha256:ef4f6578f8902c72961bcfab8c92ee33fef9f9fdc844f336667ce11422efe015

Observation 49c1450e-a640-44d9-a75a-f765678b4c7b · outbound

This paper cites T2V-OptJail: Discrete Prompt Optimization for Text-to-Video Jailbreak Attacks.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense T2V-OptJail: Discrete Prompt Optimization for Text-to-Video Jailbreak Attacks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.416164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.416164Z digest=sha256:cdf2362e7113980dc4fb17a872d0192a27c90b9302247af5eead5eeb5a8ccc87

Observation b37bf1e9-2c04-4f9a-bdb1-6d6a0da3e612 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.421711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.421711Z digest=sha256:3f6e691ecd6809a3bf15587832d3b590840dbb383cca5ec2236e3bf5552b21c2

Observation 3430bbff-3f71-4fc1-9474-13c77bfbd341 · outbound

This paper cites T2vsafetybench: Evaluating the safety of text-to-video generative models.Advances in Neural In- formation Processing Systems, 37:63858–63872, 2024.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense T2vsafetybench: Evaluating the safety of text-to-video generative models.Advances in Neural In- formation Processing Systems, 37:63858–63872, 2024

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.651236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.427081Z digest=sha256:04c5a23314be500800511c0be1031523e4e1f7b5b325bdb07a9b70cc9d59d5d6

Observation 9cbb6cf0-a596-4ec5-808f-38138c7904fb · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.432312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.432312Z digest=sha256:19bae6807df4aebb7814b25fb842049e1668ae9d5cc4dd055c1a85cef768e798

Observation c6ebe9aa-8a0d-474c-b5f6-506e679252c8 · outbound

This paper cites T2veval: Benchmark dataset and objective evaluation method for t2v-generated videos.Displays, page 103178,.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense T2veval: Benchmark dataset and objective evaluation method for t2v-generated videos.Displays, page 103178,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.634346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.439770Z digest=sha256:29190a054b13de36d45819f69e70983ffd7b77ca25d7ff234f856e8cd23be621

Observation 9085880b-0bed-4c4d-871b-ad881d48e851 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.446099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.446099Z digest=sha256:f1a48f62297152fbc6c8ff6692c0a1ff431c83e1501af54ae9c2841a95fe959e

Observation 91185768-7cf0-4da1-a1a5-4580ab734650 · outbound

This paper cites Kling-omni technical report, 2025.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Kling-omni technical report, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.618604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.451995Z digest=sha256:75397fc9efb5a01cb0ee07f49dc16ad2e85addd206e2d3577e4d47c7ae8b2637

Observation 787ac11e-2afb-4b10-959c-eebed3a757ad · outbound

This paper cites Videotetris: Towards compositional text-to-video generation.Advances in Neural Information Processing Sys- tems, 37:29489–29513, 2024.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Videotetris: Towards compositional text-to-video generation.Advances in Neural Information Processing Sys- tems, 37:29489–29513, 2024

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.602757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.457213Z digest=sha256:f397aa6995e8eef6ae154893afc72fac502ffb48a03e328022249babd2b16016

Observation aa6d094c-6ecd-415f-a8a3-05a3584c3702 · outbound

This paper cites Manipulating Multimodal Agents via Cross-Modal Prompt Injection.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Manipulating Multimodal Agents via Cross-Modal Prompt Injection

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.462806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.462806Z digest=sha256:859b6043b7da76e499b1f1aa933843e1f646d22db37f0f6217fcef1f240e88f4

Observation 7f40b8f3-844b-41c1-9968-e5b0f8480707 · outbound

This paper cites Jail- broken: How does llm safety training fail?Advances in neu- ral information processing systems, 36:80079–80110, 2023.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Jail- broken: How does llm safety training fail?Advances in neu- ral information processing systems, 36:80079–80110, 2023

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.583597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.470170Z digest=sha256:f3b552e1bea74b1bb241ccee22cf654090e303083e630c2916061d473e6ad170

Observation fc4eeec0-dd76-452b-abce-c1d7f2013989 · outbound

This paper cites Defensive prompt patch: A robust and generalizable defense of large language models against jailbreak attacks.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Defensive prompt patch: A robust and generalizable defense of large language models against jailbreak attacks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.560545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.484064Z digest=sha256:200dd3df03da8ecc706b0788a6b2164187bbaf26db55d0fcd080c1b363b50260

Observation 532d45e4-ee0e-4995-a0ca-a53616ba061a · outbound

This paper cites Video- eraser: Concept erasure in text-to-video diffusion models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Video- eraser: Concept erasure in text-to-video diffusion models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.541646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.490263Z digest=sha256:4f29a24053b36d0de59ce8057f69380e625f92baa5f8fc13d0ddfdcba8f711bd

Observation 6b4a4a8d-21e5-42fe-a6fa-0d08d8f668dd · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.500641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.500641Z digest=sha256:841ad255fb3ee20c760efe625a1bcb52376a8058635ac2d0c4acc0f208ca0fe0

Observation 81a610f3-d5d9-4ea5-a11e-e2d4588ba948 · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.509428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.509428Z digest=sha256:213765db4a10dc64a356353a6c8ce02e88bed87f1ecc58a3e4a57e3065514e87

Observation 62a1afb1-8ef7-4f5d-9bd2-81bbc9236c81 · outbound

This paper cites Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.516193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.516193Z digest=sha256:9026a214e120d9ad23998f6266ca5cbc337cc9d5f59b256ad3ab2d84b7c10edd

Observation f9a84e0f-5891-4928-8125-6b618b3d553f · outbound

This paper cites Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.521564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.521564Z digest=sha256:ab74ca53493373c5c206d70b26d91c1aa5e1195654167a3b131a7090c562bf7c

Observation dd8eaec9-7540-4b66-b0ff-cf4653a68eb6 · outbound

This paper cites Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.527677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.527677Z digest=sha256:94301c38c2b55797e01e1389c510e32255e84785cb466de680a0b4809b6dd02f

Observation d33333d5-1b61-4c26-b77e-7a84afc3ae10 · outbound

This paper cites SAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense SAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.533067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.533067Z digest=sha256:04709470ccd631ff0eb6a4b795a62c2480345d0c00e3797da2155b4a113cbbd7

Observation 2ab41e49-4e0a-42f6-bcc1-bdcd5fa9660a · outbound

This paper cites Safree: Training-free and adaptive guard for safe text-to-image and video generation.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Safree: Training-free and adaptive guard for safe text-to-image and video generation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.522597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.539171Z digest=sha256:adf73f5a57656a722e643389d5200c50d8c85530f7701c96efb265527f05205d

Observation c100e2ec-4c37-49f1-bdfa-be661290bb15 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.549719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.549719Z digest=sha256:794c623dbabba44e351d15b89a195a706a22c7c1730896aec9302ad14d00643d

Observation 19624e28-881f-42bb-964d-0cfc1773be5d · outbound

This paper cites BadRobot: Jailbreaking Embodied LLM Agents in the Physical World.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense BadRobot: Jailbreaking Embodied LLM Agents in the Physical World

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.556081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.556081Z digest=sha256:94eda73f7f3d2c59eb6f56e9b8643fec558b460d66d988dc2a334a817e349b2d

Observation 4e47367e-6738-42a7-9a3e-be9d00ec226c · outbound

This paper cites Jbshield: Defending large lan- guage models from jailbreak attacks through activated con- cept analysis and manipulation.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Jbshield: Defending large lan- guage models from jailbreak attacks through activated con- cept analysis and manipulation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:01:42.503946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:01:41.562372Z digest=sha256:4e6fbc735ff1b3645d517c5336c92132ff8e647f104a245dd49f200c25003282

Observation 9f387c79-6dd8-4fdd-9df1-3c138f0b56fd · outbound

This paper cites Prefix Guidance: A Steering Wheel for Large Language Models to Defend Against Jailbreak Attacks.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Prefix Guidance: A Steering Wheel for Large Language Models to Defend Against Jailbreak Attacks

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.567414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.567414Z digest=sha256:93403ee7a1ede877f5fb39cf81c0c0120967c5a826997a6994178d3ebb8a6250

Observation 1fa8f2b4-2334-4a8e-bc4d-9c924c6f477b · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Open-Sora: Democratizing Efficient Video Production for All

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.585415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.585415Z digest=sha256:056f2f12bc70351a811735d79dbb7518f7a39292a0ddcf3b75321478b2c2e9df

Observation 827e0253-8a58-4128-9d07-a5c44328da0a · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T14:01:41.590694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:01:41.590694Z digest=sha256:698e39e7f0e5b4202a59fde4288cb6ef65e8c0f014e0988c2fb36cf456a21061

Pith citing papers

No inbound Pith citation observations are available.