Pith. sign in

Paper Citation Record · LEDGER

Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2311.03348.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.03348 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T00:12:04.882237Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:50:11.090116Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 048aaf97-31c3-4526-bfc8-a3773f238a7e · inbound

Jailbreaking Black Box Large Language Models in Twenty Queries cites this paper.

Jailbreaking Black Box Large Language Models in Twenty Queries Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T09:48:33.242823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T09:48:31.721745Z digest=sha256:290245ff0facc1fc7c97813ce6f92dc951ef6303db852f62af3346361f52c2bb

Observation 84096182-e022-4442-8f53-f1c2d080ab13 · inbound

Dr. Jekyll and Mr. Hyde: Two Faces of LLMs cites this paper.

Dr. Jekyll and Mr. Hyde: Two Faces of LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-24T05:06:00.350340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T05:05:35.726118Z digest=sha256:21126a3554b917efd3110eadfe8ab2c901a0b7468e7cd9bd67b18c14d2357e03

Observation 49b0f1a6-1594-49cc-b309-7617cf3e94c1 · inbound

A StrongREJECT for Empty Jailbreaks cites this paper.

A StrongREJECT for Empty Jailbreaks Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:28:02.861882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T21:28:02.745230Z digest=sha256:de4236886a380d1562545ef098e013d1055a01ff7d4ef13d7d4c548ea4600bcb

Observation dc4002bc-dea2-4b34-b549-7dd66ece4c54 · inbound

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models cites this paper.

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:08:05.480000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T06:08:05.386345Z digest=sha256:3fab8988de83d31c43ef70a7986e9a6f7a5d52b6a973121eb69f437fd636eca8

Observation db505cbd-6cd1-4674-8b49-730762ea7099 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 76

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:20:44.811573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:829b2b71ef5a1d7fb504b192886c1651aa04f0e4a5dbc0d6439f6d774bb904ec

Observation 941745b9-75cd-4f8c-9a7d-aa987e2725d1 · inbound

Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models cites this paper.

Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T00:12:04.882237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:12:04.882237Z digest=sha256:02bc58df0bdbb3b9cb904e0c13cc4e9652dda6f6aac8c9c1dcef9ecd761d46b2

Observation 3d848e5d-3895-4631-9e7a-b58f4573cde8 · inbound

Agents Are All You Need for LLM Unlearning cites this paper.

Agents Are All You Need for LLM Unlearning Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T19:14:53.721398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:14:53.721398Z digest=sha256:d0db593687e8bd83139c92ede19301528c997718f9f0ce741d4f296473b3bab9

Observation 836769b0-789b-45a2-aa2e-1c59c43e35f4 · inbound

Jailbreaking with Universal Multi-Prompts cites this paper.

Jailbreaking with Universal Multi-Prompts Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T16:28:54.094781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:28:54.094781Z digest=sha256:f5b191034645902f5268248daa4cb2802c60dee0c7fcc7307f75da820efeed18

Observation fc1c11dd-ec9f-45ee-8622-dd26f0a580b0 · inbound

Position: Adversarial ML for LLMs Is Not Making Any Progress cites this paper.

Position: Adversarial ML for LLMs Is Not Making Any Progress Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T12:47:21.758202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:47:21.758202Z digest=sha256:f4d20477f344ff4165644932aeb6aad3bbe6134a2a9609ed57a47d7dc9f572fc

Observation 76c2f389-29c0-4cc6-9ff1-f7d42a703277 · inbound

QueryAttack: Jailbreaking Aligned Large Language Models Using Structured Non-natural Query Language cites this paper.

QueryAttack: Jailbreaking Aligned Large Language Models Using Structured Non-natural Query Language Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T20:46:15.926339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:46:15.926339Z digest=sha256:b4dfab08183691d76fa0549d4169ffcbd6dda2c1155c174782ecb40fed95894a

Observation e4c1861b-9875-4628-86e9-e6c8780e54de · inbound

Lifelong Safety Alignment for Language Models cites this paper.

Lifelong Safety Alignment for Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:07.771929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:07.771929Z digest=sha256:8c430021cef8b65fc3bba00ed53e0eae0e7e77a75ea9204f29d392b16025cb33

Observation 995c9188-0317-4fba-a6a0-834c209a5684 · inbound

Jailbreak Distillation: Renewable Safety Benchmarking cites this paper.

Jailbreak Distillation: Renewable Safety Benchmarking Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:44.924826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:21:44.924826Z digest=sha256:bfca43faa1b5b452072f5fdca9ebe34046ce25c09d3484ff35debc86d184e89e

Observation 3c0bbdf6-817c-4bd0-a666-601b851f6a74 · inbound

Adversarial Attacks on Robotic Vision Language Action Models cites this paper.

Adversarial Attacks on Robotic Vision Language Action Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T11:11:40.530536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:11:40.530536Z digest=sha256:79e1f25488e4bec3c525ed427c330a2d6730bfa8ac69a2206c8c45920efb115a

Observation 3c351043-6e22-4ba2-8063-e415567b38df · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.883048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.883048Z digest=sha256:10df89aefd23ba7ce38336499c46c2b6d9351adf5cf1f43652a74ee4b2705bf2

Observation 4caf03b7-64e8-467e-aa4e-56422e420bb6 · inbound

SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression cites this paper.

SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:15.645320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:51:15.645320Z digest=sha256:9a1a3746cd3f86e1513508521b7a2c49f87b6421a77be8ed1547a886353c16fd

Observation 6d906230-1a4e-4b3f-b9a8-854bff233fbf · inbound

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models cites this paper.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.859123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.859123Z digest=sha256:55395a8647d7ea934597a2c6f72c75ff951f32d0f2ceae1fd15fb8efa910b851

Observation 5942f2a1-cda5-4fab-8a27-6ec6c3c2b392 · inbound

VERA: Variational Inference Framework for Jailbreaking Large Language Models cites this paper.

VERA: Variational Inference Framework for Jailbreaking Large Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:04.720473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:04.720473Z digest=sha256:599103b8dd4c09d88dfab9ff42059cb7e259c86171fdf66e1c55acf508ac1429

Observation 74e33e9a-f10b-4f6d-88cb-9af4f587b0f0 · inbound

Linearly Decoding Refused Knowledge in Aligned Language Models cites this paper.

Linearly Decoding Refused Knowledge in Aligned Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T21:27:34.985261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:27:34.985261Z digest=sha256:22574099aeeb0ee0d256dfe7d3a09588a884d01a81d5397257591349b96aa022

Observation b853babb-024f-485f-bda8-0d6e9d553ba5 · inbound

PRM-Free Security Alignment of Large Models via Red Teaming and Adversarial Training cites this paper.

PRM-Free Security Alignment of Large Models via Red Teaming and Adversarial Training Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T17:35:48.436360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:35:48.436360Z digest=sha256:cc42e84010ceb7be2cde85d154245b7158a363d8b961168ec5afdb46ba7cf71e

Observation 02860e26-ed0f-46bf-9da8-bf95f39ad648 · inbound

From Seed to Harvest: Augmenting Human Creativity with AI for Red-teaming Text-to-Image Models cites this paper.

From Seed to Harvest: Augmenting Human Creativity with AI for Red-teaming Text-to-Image Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:20.550173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:45:20.550173Z digest=sha256:e5ab61882c9988e65b45bce3fd21bdfff163a79f05960f5546a85a702bb46be5

Observation 5bb6def3-2f52-4d35-8150-66280a73bd78 · inbound

MOCHA: Are Code Language Models Robust Against Multi-Turn Malicious Coding Prompts? cites this paper.

MOCHA: Are Code Language Models Robust Against Multi-Turn Malicious Coding Prompts? Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T14:17:34.120945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:17:34.120945Z digest=sha256:af7343aad2aa9d5a7f45f5acdc33b15acd2a9ffa968e3e417d715a8ad2986162

Observation c0049bed-32c9-4638-81e1-7145f1c03b4c · inbound

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? cites this paper.

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:42.029495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:42.029495Z digest=sha256:016dd0ea53270da2bc29ecd0d5f913d6551761520ba66fb4d5b07d0f6c7a72dc

Observation 1a625e3c-4afa-4f3b-b1e3-eb81901693df · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.745657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.745657Z digest=sha256:f0b5cb600f0787c0fbd7a0cbb03f6d67a367efe9f0249aef6d9f82a57d01e4cb

Observation 25ebd5fc-712f-4be4-8ac9-6177452bcd62 · inbound

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? cites this paper.

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-05T16:00:48.913381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:00:48.913381Z digest=sha256:aeb57b409e07cd3954e1e9171856475b96e699601a6a91f00b0e4afeb9410ce9

Observation 51bae0dc-effd-4724-9804-6e3ad33b4bdb · inbound

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs cites this paper.

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:36:52.516771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T21:34:51.665401Z digest=sha256:76fc640e4a3ae993178ae08c78702df78fb0e19ff655bc34d9bb9b976f5a1a62

Observation e652e991-c4d7-451c-a0bf-7f831e253e4a · inbound

Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks cites this paper.

Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:26:43.700266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T18:25:25.751119Z digest=sha256:750bf681901a19986e16840721208910311a8b7454380ed0c8a906bdd9170877

Observation 6fcd0746-d538-4497-a298-36eeebcd2e89 · inbound

GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models cites this paper.

GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:08:08.260439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T17:04:53.839154Z digest=sha256:e1f292059eaf17068388ef588cfa9986c8febe645667def7771cbabe5410b598

Observation b1f54b27-8df9-45c9-bd09-5e105446ab8f · inbound

Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs cites this paper.

Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:27:23.199273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T05:22:59.050348Z digest=sha256:64e5f9a2fa5e6d0000f2558cace6aba889d5a938ddbd4fe2cb3e2e099a32ea28

Observation 8cb8c199-c5ba-41cb-8c8e-6b0e44334727 · inbound

State-Dependent Safety Failures in Multi-Turn Language Model Interaction cites this paper.

State-Dependent Safety Failures in Multi-Turn Language Model Interaction Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:27.401675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:27.401675Z digest=sha256:109785f15c2c9775113ff91331ed78f755a0dec435f911f58bc832b64874c079

Observation 56730341-2f6b-4825-99bc-171d507ae8ff · inbound

Conflicts Make Large Reasoning Models Vulnerable to Attacks cites this paper.

Conflicts Make Large Reasoning Models Vulnerable to Attacks Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:51:46.135021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:24:25.255602Z digest=sha256:9fea77f0cee2d10bff5eac16d3dd718663d6ff25370623dfb9fab52c27c706f5

Observation 7f80a41d-4886-426a-9dd5-1acd3c6d89a2 · inbound

Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models cites this paper.

Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:21:00.947468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T16:05:09.033412Z digest=sha256:d48530f76931bc0df560e9d44b1bef7c405258466b9b69d6447d83213520e692

Observation 92b5cdac-c2c9-4c7f-a753-0cf32915931b · inbound

VoxSafeBench: Not Just What Is Said, but Who, How, and Where cites this paper.

VoxSafeBench: Not Just What Is Said, but Who, How, and Where Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:24:21.983100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T10:19:28.041282Z digest=sha256:c4906ac93b9cc7391518e164455bb643c515da77a320001bff32bf34e5d572c6

Observation 0cf46a83-bcf0-4960-99a5-9c5e20c551b4 · inbound

Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment cites this paper.

Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:21:07.539243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T17:24:54.796037Z digest=sha256:402a07d729684ddb431849256014d10ae3a40b11a4a64968ce556919d479ca07

Observation 429303f9-8b9c-40b1-bb68-8dd7a4bfb69a · inbound

On the Hardness of Junking LLMs cites this paper.

On the Hardness of Junking LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:21:10.199160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T17:38:32.028947Z digest=sha256:fb369211bbc379411d7bd89357a0142c54c61c3bebd550f8b6315a0281f7fdec

Observation 02566d56-245c-4f38-abd7-e26852bcab41 · inbound

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions cites this paper.

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:34:48.461754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T15:17:37.904831Z digest=sha256:7887648ce672d01e156f5354faaf8b43dc60f6e5d1c67fec1a91eafa49bcaaed

Observation 0fe7bbb2-338d-4b87-bf0e-a68e330e07ff · inbound

Quality-Diversity Evolution for Discovering Diverse Vulnerabilities in LLM Safety cites this paper.

Quality-Diversity Evolution for Discovering Diverse Vulnerabilities in LLM Safety Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-28T20:32:37.942834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T18:33:58.265424Z digest=sha256:d9d45aa941a33788d2205d38305794fe0b3ad743c4f7a82c890a9b89ce2381db

Observation ad43cf6e-9b85-4c0a-8d00-384d15da4c5e · inbound

CHASE: Adversarial Red-Blue Teaming for Improving LLM Safety using Reinforcement Learning cites this paper.

CHASE: Adversarial Red-Blue Teaming for Improving LLM Safety using Reinforcement Learning Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:06:55.661342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T02:34:26.334078Z digest=sha256:c15f93b8e62c4dbfc12d77cacde37db45c5ce042b9208cd3a89f0482a9debab3

Observation 7e0f354e-5d7b-471a-b39d-e07f8a5b6c28 · inbound

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks cites this paper.

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:26:59.288656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T01:16:07.252429Z digest=sha256:c9e95cb5291788daf15ee8c8458b9d7f3ab6301778f1aa0e19c11702f9580e67

Observation ee858e8a-052b-4b27-921c-038480946b72 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:50:11.091665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-25T20:58:53.119386Z digest=sha256:18b9237e53805d88f12b0cff57f08dfb25160cfa39328160b46aa65c8ba90b5d

Observation 3f5ce316-7ac9-4e84-bc73-956f6b4b9924 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T10:16:38.894699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:16:38.894699Z digest=sha256:b259c87582e67545ede380cb3197f2332ea5190dceedf0d245b80d7b4b8bc12b

Observation 54147477-564f-43bc-b241-f968bbf1548c · inbound

Cognitive Firewall: A Proactive, Zero-Trust, Multi-Gate Framework for LLM Safety cites this paper.

Cognitive Firewall: A Proactive, Zero-Trust, Multi-Gate Framework for LLM Safety Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:38:54.949070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T20:38:16.610310Z digest=sha256:ebe0321eee0327c824d8cc829a466d188d281cf3e2d75e840430d1106a71f8bf

Observation 10a14e1f-2063-4e44-93e4-a61675755567 · inbound

A Scalable Approach to Evaluating Moral Sensitivity in LLMs cites this paper.

A Scalable Approach to Evaluating Moral Sensitivity in LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 157

Resolution
unresolved
no resolver link, observed 2026-07-12T05:44:33.099337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T05:44:33.099337Z digest=sha256:aa43050f25cdc04107c46b80f93c4e4b57659adff0270d757b9e7989f556e055

Observation 0ae2cb66-26b5-46ae-99e1-d23cfab18d5b · inbound

Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines cites this paper.

Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T12:39:56.982241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:39:56.982241Z digest=sha256:7f6a8d5225dec9dc68b97bbd0384a68c231349832987b46f8ca0a3f279a9dc25

Observation 39690abe-ff05-411e-9ae9-0698e268e704 · inbound

Role Steering of Language Models for Social Simulations cites this paper.

Role Steering of Language Models for Social Simulations Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T01:44:31.810386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T01:44:31.810386Z digest=sha256:e107e786325e047fa08a2c096e81dc0b3840879016f5a77a913fe24ed6ca5bcd