Pith. sign in

Paper Citation Record · LEDGER

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs

As of 22 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 3 inbound Pith citation observations for arXiv:2412.15623.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.15623 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T11:17:49.131015Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:46:19.440645Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T07:13:16.297937Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6af88e4b-0463-4587-9adb-a98e97fcb927 · outbound

This paper cites GPT-4 Technical Report.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.932972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.932972Z digest=sha256:490ba4eba080f0c30dcae07b17d3da4dcb1e9fe0a125d502fa092f58b5335523

Observation 12f2a6eb-f3da-48a6-814c-2d6b4a329066 · outbound

This paper cites Detecting Language Model Attacks with Perplexity.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Detecting Language Model Attacks with Perplexity

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.938143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.938143Z digest=sha256:396a9464bda46171b0130bd5e004fefb1bdad576842d508e5e25541f011aae7e

Observation 5628e92e-457a-4afb-a72f-f78b17f5a140 · outbound

This paper cites G.; Guo, Z.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs G.; Guo, Z

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.943269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.943269Z digest=sha256:7c5ec100f42de436a0e9d39c5a68ab115109105749dd59f4d6681b79cbd16957

Observation 2acfd132-0dc6-4439-93ae-e98856297630 · outbound

This paper cites Qwen Technical Report.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.948353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.948353Z digest=sha256:59e231ff9771bf6e0a1ecd1b3bb2ad57ca94d61521951391f246ab76f1efd48d

Observation 168bfd0a-316d-4120-8615-4698b20260c2 · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs On the Opportunities and Risks of Foundation Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.953494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.953494Z digest=sha256:5bd013dd42f44d29b1ef494093123bb72397c5a213ae2ba3a2c10ab2b4868ec1

Observation 238f35a9-ef22-4939-b519-220aec9528ec · outbound

This paper cites A.; and Terry, M.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs A.; and Terry, M

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:17:49.702454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:48.958936Z digest=sha256:cba8f82271e8a767ab1ffaf6afbb848fd4dc6cf397e944a7ac235366ade2e2e4

Observation aac5db27-5d32-468f-8b23-b5d8209259b5 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.963693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.963693Z digest=sha256:e393cca656fb42c91f0aef40bf59247c51499ad2915400fe80bbd1c40cea5986

Observation 469e2278-cd99-4339-b0e1-c1d0e748ed52 · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.690333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:48.969079Z digest=sha256:2e2181f1a7e0a27957731f256c177eec3a22dc508dfbbf8188f168b15f3bd3c2

Observation 728c2b89-3e74-416b-8fa6-cbac83fce0f3 · outbound

This paper cites Multilingual Jailbreak Challenges in Large Language Models.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Multilingual Jailbreak Challenges in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.973813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.973813Z digest=sha256:f9b293538e9e07600470aa5cb1919e92c755eef9bd2ac349f1ffc2825b300090

Observation 5865d1d9-33db-4f74-a8c9-2a169c4fbcee · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs KTO: Model Alignment as Prospect Theoretic Optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.979129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.979129Z digest=sha256:434229a2e61a927d95e1a5375147126df8c129edc898d2b50423b484d2e4f7f7

Observation a6801a13-e451-4995-91f1-611a40a04a97 · outbound

This paper cites A.; Matar, N.; Sowan, B.; Al Khaldy, M.; and Barham, H.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs A.; Matar, N.; Sowan, B.; Al Khaldy, M.; and Barham, H

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:17:49.679038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:48.984219Z digest=sha256:505a22f733fc32e173ad5504e36ae3e05d8f4498241d74e2bcaba0113ea4a5aa

Observation 42b34650-d4e8-4019-a009-6fc9e15caba8 · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.666583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:48.988512Z digest=sha256:63868648b6f38380002872566a2e12c6186e37eb944b5ac3485341b013689216

Observation e877a6fc-77ba-44a0-8685-0b97b9679712 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.992980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.992980Z digest=sha256:422ea34db41bc2f9d16d86e673f53e435bce18be77b4f5407066e94a235d5be0

Observation 24ebf4fe-5310-4ecc-a900-d493c13b42e1 · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:48.997434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:48.997434Z digest=sha256:aa0ae080c3e3b154861cee048ec79e3a66debeec80d3f358a0abe6807c5677cf

Observation fc0e9486-c296-420f-8cec-e4bce0a2f699 · outbound

This paper cites AI Alignment: A Comprehensive Survey.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs AI Alignment: A Comprehensive Survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.002361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.002361Z digest=sha256:88099e344cf6f30e7545b3644866ae8504a12b03a259be843bdb24eb579ee0c1

Observation 0557ef9e-4c06-4bf8-9bc8-2687cd2009fc · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.655666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:49.006880Z digest=sha256:eb24b26975d55590691ce41f36800ab70660c8d2596eda43696a421ef3267e36

Observation c03a3bdc-2396-42e6-a34d-6ab8f9af59d4 · outbound

This paper cites Multi-step Jailbreaking Privacy Attacks on ChatGPT.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Multi-step Jailbreaking Privacy Attacks on ChatGPT

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.011649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.011649Z digest=sha256:65b8b39490da243427d17b23ffa6b22e76611836667570db61a52f73429c155c

Observation 824633dc-76df-4737-ac19-52cdbd47e167 · outbound

This paper cites DeepInception: Hypnotize Large Language Model to Be Jailbreaker.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs DeepInception: Hypnotize Large Language Model to Be Jailbreaker

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.016707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.016707Z digest=sha256:c3c7fc10dd595b8d1a2645d63aa626fd8ac1e1e2215cbc0ef12b7096576ed6e8

Observation 0392e662-461c-41e5-bbe9-a5b16b01306f · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.644753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:49.022641Z digest=sha256:c33f23fff2dfa34cf3a673c5879acd42519df47bc8311eca7d82c5345ada6bd8

Observation 9293c9b8-8dee-47e2-9b81-0a3d9ae76639 · outbound

This paper cites Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.027329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.027329Z digest=sha256:4f6ea735a1c9352e9865d51155b407e7063fd5013c6824e730e296f451bac984

Observation 7e34296a-7701-469c-a51c-ec0dffb8eec0 · outbound

This paper cites SimPO: Simple Preference Optimization with a Reference-Free Reward.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs SimPO: Simple Preference Optimization with a Reference-Free Reward

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.032269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.032269Z digest=sha256:1cc4565e46f5893042511d8c22dfa246dc0ea7c71f655c27942bae2ed8e75f2a

Observation 01b7bb61-1a87-4d64-a81d-178a395017dc · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.037916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.037916Z digest=sha256:33fca14d66b91dd211bfd4066545c1bd21cbe7f40e9838fe71b067b8a8381519

Observation 0f58b3d2-fce9-49ad-b22e-fbb22562a0aa · outbound

This paper cites Red Teaming Language Models with Language Models.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Red Teaming Language Models with Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.043581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.043581Z digest=sha256:65375e7ae4da42bb4a1d0105a10ad54b17dc964732613dde29f7df63a3e17faf

Observation 78dacdc9-dbe9-4076-8d02-912a023e24ca · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.626497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:49.048142Z digest=sha256:4eac3098cd7aa68ae34ce6cc14e3707a4056e9ae876829027cbd511c5c75d697

Observation 9803a7e5-5749-453b-bc7c-098d46ebf5bf · outbound

This paper cites D.; Ermon, S.; and Finn, C.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs D.; Ermon, S.; and Finn, C

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.052422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.052422Z digest=sha256:4373d0c9e27e7b92b9f7ba80e0422a6b00e049e72dee80be4bc90622c2d6aab1

Observation 8e748bb9-50f7-42ba-8c86-6d59fd1dd289 · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.607845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:49.056580Z digest=sha256:5692b6577c32e9cdc59934db0f5b131810e999970b091f2a11fc5d6fb2359db5

Observation eb2dddc6-044a-4eee-a744-3e8134cf1f41 · outbound

This paper cites Large Language Model Alignment: A Survey.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Large Language Model Alignment: A Survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.061766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.061766Z digest=sha256:f88771ce3f69708c93d8da98951bdd7497e89e2abd7ee2a1b919c98d99ae31b3

Observation 6245f067-4c2b-4693-89de-9ae3a6ff5ba4 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.069826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.069826Z digest=sha256:5a4a7e64737cee2c3cf1917a44263a89f5d01eab82e391604e9493f8a77a62cb

Observation f4e4fe56-7535-424b-b328-10d7df17daf0 · outbound

This paper cites Aligning Large Language Models with Human: A Survey.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Aligning Large Language Models with Human: A Survey

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.074644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.074644Z digest=sha256:c36817e055642e859d393ff6252b4a639362598e10bf670c492fe750c6420fa0

Observation 4d49c7c0-02b0-416a-a604-9eb4492650b8 · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.596508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:49.079154Z digest=sha256:b4c78aa1d44bc1cf8eff9ce3a9ddd93107368b29a0b7e3f82d8fade6897f3585

Observation a36afdaf-8631-403f-b79d-4c01ef20fcce · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.585887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:49.083712Z digest=sha256:3ac8f94d46ba1c46ec92f40707bd53711b27fc5c3ec6ea224568b067dde07566

Observation bbb33e81-3881-47da-af11-20f837f38f5d · outbound

This paper cites Cognitive Overload: Jailbreaking Large Language Models with Overloaded Logical Thinking.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Cognitive Overload: Jailbreaking Large Language Models with Overloaded Logical Thinking

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.088485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.088485Z digest=sha256:19c8ea2ab42cbc9bcd25cc4f7d1cfada4315411111bc57a442ba00a0b56fe647

Observation fb530ff6-8e8f-44cf-a7f8-d49f313fa476 · outbound

This paper cites A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.093000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.093000Z digest=sha256:3c0e40b50d5ee38693318b6ffe95338d5a427bcbd16565d5015ebf0b39cdc754

Observation 916a2252-3f75-4f40-a54e-6a8d0050f7c2 · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.097099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.097099Z digest=sha256:361027fda877ad5086d04e6fdffcb26bb093b809b2e0e74ffea6c382190927e6

Observation d71b151d-cb61-47b7-81fb-bf21d2acfc46 · outbound

This paper cites an unresolved cited work.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-11T11:17:49.568554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-11T11:17:49.101446Z digest=sha256:f7019dac98ba1164ad888e64354b72f013daec1471bbfb10c615caff512f3f3d

Observation 4a5f5dfe-42d3-458a-9933-33c402fa66f5 · outbound

This paper cites Low-Resource Languages Jailbreak GPT-4.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Low-Resource Languages Jailbreak GPT-4

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.105339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.105339Z digest=sha256:6b44a6c9e3061f7a58d10a3436e8c60c7d974b50561c37bb0c76e60d6f47dff9

Observation 560d5cea-4212-4f18-9b72-ced4d181817e · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.109167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.109167Z digest=sha256:f50f1f488258d7a6c99621c6de78625d7403a44f925b7a04bbdfbe4bb6ed0689

Observation 137eadbd-56b4-4846-a67d-8e4022145685 · outbound

This paper cites GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.113067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.113067Z digest=sha256:1659a5c82c4bf4f845640e7a3864dc7a74408644742f167d7535f542fc254aa3

Observation ba866eeb-24be-4a4f-ad06-002dc1a6d468 · outbound

This paper cites JADE: A Linguistics-based Safety Evaluation Platform for Large Language Models.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs JADE: A Linguistics-based Safety Evaluation Platform for Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.117092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.117092Z digest=sha256:888f7b2aefafde7f428ccc4f75b5bb224def2f28f7016101411bd39695283fb1

Observation d0a73ea6-7be6-4e3f-9b3c-79e9c34d6bbc · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.121982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.121982Z digest=sha256:dcec62923e7efd1fb073773b27f465d639648754dba0e501b80fa06ccad2b9d9

Observation 004e798d-c51f-4a46-8c99-fd46fe2f21c1 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs , " * write output.state after.block = add.period write newline

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.126570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.126570Z digest=sha256:1ffdd8c80e5955ed5ec342010a687215052a773189f00f795ac131932743c4da

Observation 7738d14f-ae98-41a4-87c8-f2e914c06812 · outbound

This paper cites write newline.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs write newline

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.131015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.131015Z digest=sha256:ab116618f574e16d099d5dcf6c6576cbec2a41d3ec39f34d3006d8839c90252c

Pith citing papers

Observation bb2d9b8e-3122-4bbf-bf55-803385c302ec · inbound

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems cites this paper.

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs

Reference 175

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:19.440645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:19.440645Z digest=sha256:5230499c1240efb72812c6176d525ba443883a16243ec48505bc982ea350c865

Observation 050f4c5c-0847-4f91-a0ba-2f68a0cbc320 · inbound

Safety Alignment of LMs via Non-cooperative Games cites this paper.

Safety Alignment of LMs via Non-cooperative Games JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T14:24:21.893602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T14:24:21.893602Z digest=sha256:19081bd5466588548f5defdd5d8931101f845b1759c26f4a7d3900c7c9500f04

Observation 0bb914fe-945d-4138-8a3f-3b35c25e4e5b · inbound

Evolving Skill-Structured Attack Memory Enhances LLM Jailbreaking cites this paper.

Evolving Skill-Structured Attack Memory Enhances LLM Jailbreaking JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:13:16.299680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:10:50.007951Z digest=sha256:adeef473703192c1324b9607c27d3aec2a81c747931538925910c2db39268a2b