Pith. sign in

Paper Citation Record · LEDGER

The bitter lesson of misuse detection

As of 7 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2507.06282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.06282 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:18:40.850181Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2c91e3e0-d55e-4ba2-af12-aebd325fadf1 · outbound

This paper cites GuardBench: A Large-Scale Benchmark for Guardrail Models.

The bitter lesson of misuse detection GuardBench: A Large-Scale Benchmark for Guardrail Models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.818607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:38.633982Z digest=sha256:59acf0a3585c31e0451802c77967440701795d494a08a3814e001674a2ee0cb2

Observation a52d6d52-2bd2-43c6-b438-7005ac68bff6 · outbound

This paper cites Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming,.

The bitter lesson of misuse detection Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.563105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:38.761731Z digest=sha256:dc7c0ed33826585a872470c8a896279fa72a2262457591a67c6c20cf0743bc55

Observation f8c82631-8dd1-4ee7-8b64-f927f3aa1a4a · outbound

This paper cites NeurIPS 2024 Datasets and Benchmarks Track, 2024.

The bitter lesson of misuse detection NeurIPS 2024 Datasets and Benchmarks Track, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.272682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:38.810817Z digest=sha256:ee5b867717de958985e220727ded0affa492d167b5e6f7038ff53159f624d974

Observation 039fb256-65d4-4e2c-9e55-e0bbc9b68aee · outbound

This paper cites "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models.

The bitter lesson of misuse detection "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:38.885620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:38.885620Z digest=sha256:f09d2f08b58ca0dbb9a6a28a439ae020eb69b29133485a671c326ec55d4cf8da

Observation 52615106-1ebe-4879-9046-de714e9e8211 · outbound

This paper cites SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal.

The bitter lesson of misuse detection SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.027078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.027078Z digest=sha256:54ff4754f141a050d5b60940c6c7e12e05a07bb6c361cd7a7631a15483d58a26

Observation df1cfef7-9e07-4e97-95c8-dabb2346ac2e · outbound

This paper cites AdvBench: Universal and Transferable Ad- versarial Attacks on Aligned Language Models,.

The bitter lesson of misuse detection AdvBench: Universal and Transferable Ad- versarial Attacks on Aligned Language Models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.058381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:39.230798Z digest=sha256:45dc498e46f6c05bcc0adacfe85974c0ceb61a057aadf210de1b3ef5b1dfcb3d

Observation 23662cfe-a460-48e7-96b9-ee0dac2f98c2 · outbound

This paper cites CatQA: A Dataset for Categorizing Questions as Safe or Unsafe,.

The bitter lesson of misuse detection CatQA: A Dataset for Categorizing Questions as Safe or Unsafe,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.891885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:39.311873Z digest=sha256:c650a128a350d2f40b9a0a6c3800d9f2e7bec9c3985ed9a69f6e02e081c294fe

Observation f1c918c9-da12-4d60-8c4e-c6abb5eddbd7 · outbound

This paper cites Do Not Answer: Testing AI Refusal to Unsafe Questions,.

The bitter lesson of misuse detection Do Not Answer: Testing AI Refusal to Unsafe Questions,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.701691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:39.365817Z digest=sha256:70af0ad42791af06b7abcfe5bf9cd6d2c69bfca9775da4e3a17d3c55d6899d32

Observation c0105b1f-cfac-43dd-8207-35d7d779c5c9 · outbound

This paper cites HH-RLHF: Helpful and Harmless Reinforcement Learning from Hu- man Feedback,.

The bitter lesson of misuse detection HH-RLHF: Helpful and Harmless Reinforcement Learning from Hu- man Feedback,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.506895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:39.472725Z digest=sha256:f7c7fbbb58b1835b0e0b593054494d0d52c93aea88de3dae14504b2b9b5f21c1

Observation 8543b09b-adba-4383-be52-c9d1d4dff676 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

The bitter lesson of misuse detection HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.588048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.588048Z digest=sha256:921306f8705e67efe38715e3ba0e1206ab25a838970188ecd9f1e1ac193aa175

Observation 9a3352f7-c722-4e0a-92c8-3a166a2de097 · outbound

This paper cites A StrongREJECT for Empty Jailbreaks.

The bitter lesson of misuse detection A StrongREJECT for Empty Jailbreaks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.703539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.703539Z digest=sha256:9e6cc024d91bddd013c600ebc7584c357df728d1491ae3d41032692f8b55290d

Observation f6b9e1a3-e909-4b4a-9aab-9f0e404c9aac · outbound

This paper cites The Twelfth International Conference on Learning Representations, 2024.

The bitter lesson of misuse detection The Twelfth International Conference on Learning Representations, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.239110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:39.798982Z digest=sha256:365a7685fb5afedc8962e848726e91678b4b0fcf0bc2782fee4b2d3f30ec0f25

Observation 82e96a74-af9e-44c2-99d9-b2a6927de0eb · outbound

This paper cites an unresolved cited work.

The bitter lesson of misuse detection Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:18:42.037851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:39.887601Z digest=sha256:c06e608f3b42d2d01bf5f01544b84e6bd19a4d5b9362ed94758c28577821bac4

Observation 219f32df-033c-482f-bf49-dc911055af8f · outbound

This paper cites an unresolved cited work.

The bitter lesson of misuse detection Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:18:41.743137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:40.022590Z digest=sha256:f58a08848ec8c2e9643dcafcd679f53691d5459451d8a807bdcdb3d756a03e4e

Observation 7efef56f-9ccc-4f3e-b0c4-589bac84fd58 · outbound

This paper cites DeepInception: Hypnotize Large Language Model to Be Jailbreaker.

The bitter lesson of misuse detection DeepInception: Hypnotize Large Language Model to Be Jailbreaker

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.154024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.154024Z digest=sha256:bdcaaca115f1757fba9ee550ced6bde04f70dd4b1d76e42293fc6e5581ee8600

Observation 05f8fcb2-eb3f-441e-99f1-8a526aa9d2fe · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

The bitter lesson of misuse detection Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.283148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.283148Z digest=sha256:10196f25669a48eaaae9301b0d09a00f474ae1032b8f1bc5c54c4ac429da1dd4

Observation 25f66079-b23c-41d8-b039-4b95459d3b93 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

The bitter lesson of misuse detection Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.429244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.429244Z digest=sha256:48b472236087862dabe53ba2c1334af81a587fa250103226bd717f692babd933

Observation cdba531b-0cde-4c13-b6f1-c526f77578c9 · outbound

This paper cites AI Control: Improving Safety Despite Intentional Subversion.

The bitter lesson of misuse detection AI Control: Improving Safety Despite Intentional Subversion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.580182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.580182Z digest=sha256:04cdf13e4892ed1ffe230df8f44fcc00b516bc05e19576ece3babae5ce3af3a7

Observation a53b041b-1dfd-4607-9626-78256768fe15 · outbound

This paper cites Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?.

The bitter lesson of misuse detection Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.663885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.663885Z digest=sha256:c8039b9ca7a70566e9167d7e0aade234c1ddc7c561fb3d7510cd73225ef4f78a

Observation 15b373f1-d82b-4fc5-b86b-a2de051580f3 · outbound

This paper cites Is this prompt harmful or not?.

The bitter lesson of misuse detection Is this prompt harmful or not?

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:41.442675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:18:40.850181Z digest=sha256:57063fe3c640050c3d3b562ba25a9ecfc218efbea224a70c11e0db1359b30f8c

Pith citing papers

No inbound Pith citation observations are available.