Pith. sign in

Paper Citation Record · LEDGER

Best-of-N Jailbreaking

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2412.03556.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.03556 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 39 of 39 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:02:47.243677Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:30:00.621770Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 85ec56fe-01e7-40cf-b3fe-6caa0440d7be · inbound

Jailbreaking to Jailbreak cites this paper.

Jailbreaking to Jailbreak Best-of-N Jailbreaking

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T17:02:47.243677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:02:47.243677Z digest=sha256:7265ea3d71bae7c04d8160d38cc6e9c58d1cbd8f4dfb7aaa7a3b59fd576a3f98

Observation 439c5113-cf97-4dc3-9efb-3ed1a5097c5c · inbound

LLM-Safety Evaluations Lack Robustness cites this paper.

LLM-Safety Evaluations Lack Robustness Best-of-N Jailbreaking

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-23T01:27:21.325650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T01:26:45.402983Z digest=sha256:c1c46f50e1c3148bd960333ba13756d2ba64e0d0001f8fe0ead89b935b2febfb

Observation 04786a40-3de7-4e1c-9843-f758d418ff67 · inbound

Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMs cites this paper.

Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMs Best-of-N Jailbreaking

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:41:41.482545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T14:40:58.506345Z digest=sha256:4cf9e30c54675465ec265234e11e9c75e3ab7e012ac0b33f2a6627622dd82eee

Observation 2eb2895a-8b70-4ba9-a4b1-72f283573b16 · inbound

Audio Jailbreak: An Open Comprehensive Benchmark for Jailbreaking Large Audio-Language Models cites this paper.

Audio Jailbreak: An Open Comprehensive Benchmark for Jailbreaking Large Audio-Language Models Best-of-N Jailbreaking

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:23:01.792126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:23:01.792126Z digest=sha256:ecc50fcec607d06370e39688a95054bcf06323959f882abd20bee8f26dbd68e4

Observation b82b1bb9-91ea-4b2f-9b36-c9859225bbc6 · inbound

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey cites this paper.

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey Best-of-N Jailbreaking

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T13:34:53.247293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T13:32:57.771753Z digest=sha256:5e040612f50f3f2b09f71d7408917c0eec1eaefb8f77b3cdb811dde61188a1fe

Observation 32f2bbe1-000b-499e-944a-72cae53d7fe7 · inbound

An Example Safety Case for Safeguards Against Misuse cites this paper.

An Example Safety Case for Safeguards Against Misuse Best-of-N Jailbreaking

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.930783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.930783Z digest=sha256:7fc7e1ee7d277bcf54cb48b4229aaaf0f2335902cafd11aa30acc709c960adfb

Observation ca9c6f7a-6ddd-4ef9-9e2c-e833b8540891 · inbound

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning cites this paper.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Best-of-N Jailbreaking

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.712974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.712974Z digest=sha256:43f18ff776a797a13c933de7793b14895aeb5ff677d98498d5dd4a878d2d010d

Observation 97d18df2-a1bc-4810-bbb2-5b67da0a24f9 · inbound

A Red Teaming Roadmap Towards System-Level Safety cites this paper.

A Red Teaming Roadmap Towards System-Level Safety Best-of-N Jailbreaking

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:19.700239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:19.700239Z digest=sha256:27d4b38d572e5c16f712e1dda1f04aa372da743fdefa749aa1238fba1e6eac09

Observation 2cb4d93b-ad9b-4a88-a555-149af2568ed9 · inbound

JavelinGuard: Low-Cost Transformer Architectures for LLM Security cites this paper.

JavelinGuard: Low-Cost Transformer Architectures for LLM Security Best-of-N Jailbreaking

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:25.408036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:41:25.408036Z digest=sha256:7a8c6ed1ff1e0e9709ee0106c5b95e27242a28e22277d5dfa293fd629256f9c7

Observation f0dc3f90-d83f-4cc1-860c-a87f36c1f26d · inbound

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models cites this paper.

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models Best-of-N Jailbreaking

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:44.141461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:08:44.141461Z digest=sha256:4bf422d97ebc27dc82207c59a059ee9d3f93feff11b338425350869c2972b6a9

Observation 09d6e90e-bba0-4890-9cb1-187c5f4cdcb6 · inbound

Min-p, Max Exaggeration: A Critical Analysis of Min-p Sampling in Language Models cites this paper.

Min-p, Max Exaggeration: A Critical Analysis of Min-p Sampling in Language Models Best-of-N Jailbreaking

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:55.835946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:31:55.835946Z digest=sha256:8f29e761c750e55abc18d3ec61b00d9c690f4305fe72cc0d4f9b974ca52d5c8d

Observation 3a6207c8-1ab4-4775-818f-960bb334f4d5 · inbound

Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem cites this paper.

Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem Best-of-N Jailbreaking

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:42:12.775051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T08:40:56.186349Z digest=sha256:4691ddccdad499131537b737c403949a02c961fbe62a6411697443f0ece74924

Observation 71faf483-00ac-4258-909e-e4c8a15e3999 · inbound

Best-of-N through the Smoothing Lens: KL Divergence and Regret Analysis cites this paper.

Best-of-N through the Smoothing Lens: KL Divergence and Regret Analysis Best-of-N Jailbreaking

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T19:28:55.320762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:28:55.320762Z digest=sha256:2cc42202719bf6f8811309e39278411d93bf46fa5a438db715906816b4c81e86

Observation 2b6281d1-9e5b-4266-aeac-f6ee9d57ec27 · inbound

SATORI: Static Test Oracle Generation for REST APIs cites this paper.

SATORI: Static Test Oracle Generation for REST APIs Best-of-N Jailbreaking

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T17:26:48.687153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:26:48.687153Z digest=sha256:de0c24ce62ff1b05a8670bc85206ad5ef288970644bd5bee87fb4f48d0b7b320

Observation e7405181-1c1b-4e2e-9c0a-0ca84f51fc96 · inbound

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models cites this paper.

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models Best-of-N Jailbreaking

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:42:36.859445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T12:41:48.620040Z digest=sha256:b55ba3c1376f309bad1f3fecee684d5d859a92bba4e8c10a0ce4c75a52cac36f

Observation eb98457e-6fda-4a41-9c3e-2350d4eb2a4e · inbound

MARS: Margin and Semantic-Aware Data Augmentation for Reward Modeling cites this paper.

MARS: Margin and Semantic-Aware Data Augmentation for Reward Modeling Best-of-N Jailbreaking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T22:14:32.548751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T22:14:32.548751Z digest=sha256:b055a9e1325181905cc36c394b27638d55292db887ec97786bd7ccd4c202349a

Observation 748a4f51-d459-423a-a7b7-9197d1330821 · inbound

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking cites this paper.

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking Best-of-N Jailbreaking

Reference 15

Resolution
malformed identifier
arxiv_id, observed 2026-05-11T08:21:01.222981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:40:59.993298Z digest=sha256:490ac9c327af013a95ea71dc1e21cf30c87cb0e1970e66e5afc699a1b7d90849

Observation d7c40210-65ab-4ab9-9eb3-9c7c38aafe04 · inbound

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs cites this paper.

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs Best-of-N Jailbreaking

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:21:00.039744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T16:41:52.440793Z digest=sha256:54a43f82a251b7783150f711c140914450afdb1dc4b4c3e275c48d85bbb953a7

Observation 4201f267-ff94-43bb-bb96-0ba7afb39d11 · inbound

A Synonymous Variational Perspective on the Rate-Distortion-Perception Tradeoff cites this paper.

A Synonymous Variational Perspective on the Rate-Distortion-Perception Tradeoff Best-of-N Jailbreaking

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T20:09:57.922788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:09:57.922788Z digest=sha256:290e86f3683f3ede30f6d99d7a6d2649b0639a351f21fdca70f5da599a0f6743

Observation 4bc44e3b-4f87-42ed-a879-9bd0bc9511d8 · inbound

Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection cites this paper.

Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection Best-of-N Jailbreaking

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:35:18.894364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T11:32:10.126062Z digest=sha256:0f82a21031d799fc6a04068111d3494d5ab22e2d777999da746280c5619553f5

Observation ace32fe4-dfee-431c-ac2f-40434fb6024f · inbound

Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs cites this paper.

Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs Best-of-N Jailbreaking

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:02:24.897246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T08:01:25.938248Z digest=sha256:09f1ca4ed7bca448bcd23dadb8725dac68c5109462edd700f742c1be320201ed

Observation 97849cd9-806d-4283-a5d7-26bb9ceeb6c8 · inbound

Estimating Tail Risks in Language Model Output Distributions cites this paper.

Estimating Tail Risks in Language Model Output Distributions Best-of-N Jailbreaking

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:08.195272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T12:26:41.797166Z digest=sha256:9eb52c28f791d7c396297f67ba97c3e01db20f1e908688638954e73d78d6f7c7

Observation 51287724-196a-4305-a05b-228e572319ed · inbound

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs cites this paper.

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs Best-of-N Jailbreaking

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:01:27.514254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T08:34:14.310656Z digest=sha256:3d32bd2418b555b62d96cbc6b8ad72423ed395d73f84ae708052c717991e552b

Observation 5aaa337b-cf50-42b5-ae66-09c6cf565524 · inbound

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety cites this paper.

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety Best-of-N Jailbreaking

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:31:01.204907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T16:00:32.413225Z digest=sha256:87ec99391e59f5e0d533a019c86f00cba199305d02ca5e87a3980c9e8b996248

Observation f0ccd0e6-ea9b-4fb5-b623-3bdbb7df0bd0 · inbound

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation cites this paper.

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation Best-of-N Jailbreaking

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:05:36.111722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T18:55:24.214794Z digest=sha256:e80e00462b11f21e3e0ab391a2c422bfa04774591d348bf8d93f02aa3a44844e

Observation ac000546-95df-48b8-bf80-54421d502cb7 · inbound

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation cites this paper.

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation Best-of-N Jailbreaking

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:15:09.542059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T00:06:52.820343Z digest=sha256:6f11eb3627c941f663f5d91f2906c0bc500ab3f5b77ed6940e0108973944a6a6

Observation 75ed7f92-0eb0-467c-b7cb-86c914587a7c · inbound

Exposing LLM Safety Gaps Through Mathematical Encoding:New Attacks and Systematic Analysis cites this paper.

Exposing LLM Safety Gaps Through Mathematical Encoding:New Attacks and Systematic Analysis Best-of-N Jailbreaking

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:51:31.128666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T15:48:54.277233Z digest=sha256:28658d7146629364b47534da3840ecbbd05d800192da55e82867ab3d74a57136

Observation b5d508a6-1c86-44ba-a524-64871177957b · inbound

SoK: Robustness in Large Language Models against Jailbreak Attacks cites this paper.

SoK: Robustness in Large Language Models against Jailbreak Attacks Best-of-N Jailbreaking

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:01:08.873331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T16:42:41.137808Z digest=sha256:6ae480b3cd2d04c9479019130af4be3dfc1e5258afc0c2fdc72879312fd75252

Observation b43a254e-e534-4573-bf85-065be58d59b5 · inbound

Quantifying LLM Safety Degradation Under Repeated Attacks Using Survival Analysis cites this paper.

Quantifying LLM Safety Degradation Under Repeated Attacks Using Survival Analysis Best-of-N Jailbreaking

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:12:51.266267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-14T19:10:29.055784Z digest=sha256:2859fbbba3602b5b7054b995734f7d68632f85b3cac52d28af1f3fa313248fd9

Observation 0566c502-91e6-4f09-a7d2-62fdb3339418 · inbound

The Great Pretender: A Stochasticity Problem in LLM Jailbreak cites this paper.

The Great Pretender: A Stochasticity Problem in LLM Jailbreak Best-of-N Jailbreaking

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:13:31.268689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T02:09:42.090550Z digest=sha256:c9278b15b6b5c1b607702ec5641b254ac58643e502494a11ec10d2d07818eece

Observation a8e2ceb8-9c24-4451-b9e4-2df91ca83f8a · inbound

Compositional Jailbreaking: An Empirical Analysis of Mutator Chain Interactions in Aligned LLMs cites this paper.

Compositional Jailbreaking: An Empirical Analysis of Mutator Chain Interactions in Aligned LLMs Best-of-N Jailbreaking

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:33:38.092245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T18:31:05.740757Z digest=sha256:9020816740936ad543dabd7e554b530ccd79cfca54743801ac81b2accd4e3e24

Observation 418134a3-e0f9-4bb2-a27b-42b369799dda · inbound

Babel: Jailbreaking Safety Attention via Obfuscation Distribution Optimized Sampling cites this paper.

Babel: Jailbreaking Safety Attention via Obfuscation Distribution Optimized Sampling Best-of-N Jailbreaking

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:08:11.947366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:08:07.648295Z digest=sha256:dca7258a00584cd60bcb2d852badafb458e1df866ba8a623094dedf97127e8f1

Observation 22e9386b-000a-4dc3-929e-7761e40ff97c · inbound

Jailbreak susceptibility prediction and mitigation via the behavioral geometry of models cites this paper.

Jailbreak susceptibility prediction and mitigation via the behavioral geometry of models Best-of-N Jailbreaking

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:53:47.735743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T17:43:47.849960Z digest=sha256:1bb9438cdc932e560b4994fce6ce00b3a539c67e3d07e06bab1c397f7cd1ca33

Observation 61143039-f524-45b3-9194-26db9b2a0d36 · inbound

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs cites this paper.

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs Best-of-N Jailbreaking

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:16:34.766731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T09:21:57.373862Z digest=sha256:2040058add47d1e47f4782207356fcb0049322abe1abe239a45147b597231107

Observation c793643a-f44e-44b3-81ca-4f04d32df8b2 · inbound

Item Response Scaling Laws: A Measurement Theory Approach for Efficient and Generalizable Neural Scaling Estimation cites this paper.

Item Response Scaling Laws: A Measurement Theory Approach for Efficient and Generalizable Neural Scaling Estimation Best-of-N Jailbreaking

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T22:52:45.594721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T22:48:20.193699Z digest=sha256:05f68377c4ef4c8de4a6be23d274418f26823a68657a3202481fdaf1c7b9b36e

Observation 593d994f-d8ad-46f9-91e7-665bcb1b3b3b · inbound

Do Thinking Tokens Help with Safety? cites this paper.

Do Thinking Tokens Help with Safety? Best-of-N Jailbreaking

Reference 74

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T17:30:00.623688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-25T23:37:49.412578Z digest=sha256:e6d2b97be77749e361af1a22063f967f94f791a5ea51d1f90402bc11453b2c33

Observation 519ed499-c4de-4d3c-82bd-b10b01ae74aa · inbound

Breaking Safety at the Token Boundary: How BPE Tokenization Creates Exploitable Gaps in LLM Alignment cites this paper.

Breaking Safety at the Token Boundary: How BPE Tokenization Creates Exploitable Gaps in LLM Alignment Best-of-N Jailbreaking

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:19:20.111530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-04T01:17:44.209685Z digest=sha256:eb585836e88d246428961be182db8e65aec4a9b6809520904be9a80533819aa5

Observation e3a773bf-f1be-473d-9c98-71834a7ad365 · inbound

DecompRL: Solving Harder Problems by Learning Modular Code Generation cites this paper.

DecompRL: Solving Harder Problems by Learning Modular Code Generation Best-of-N Jailbreaking

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:38:39.781863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-03T16:30:34.793328Z digest=sha256:b930690e92ab0503ee5f1ef434cb8cb9e8961ffeefa3a2424be74eaaf585585f

Observation fe14dc20-04de-4e94-9e16-600890e0727c · inbound

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity cites this paper.

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity Best-of-N Jailbreaking

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T00:48:48.102460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T00:48:48.102460Z digest=sha256:a2a227d83ddc9a3c28f3f6fa9e40e265170c214fad3ed11d89fc658f592db312