Pith. sign in

Paper Citation Record · LEDGER

The bitter lesson of misuse detection

As of 7 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2507.06282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.06282 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:18:40.850181Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2c91e3e0-d55e-4ba2-af12-aebd325fadf1 · outbound

This paper cites GuardBench: A Large-Scale Benchmark for Guardrail Models.

The bitter lesson of misuse detection GuardBench: A Large-Scale Benchmark for Guardrail Models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.818607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:38.633982Z digest=sha256:b341afb66e2621684b3fe249df598278a6b6aab50aeb4084f3da1f5295b306c9

Observation a52d6d52-2bd2-43c6-b438-7005ac68bff6 · outbound

This paper cites Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming,.

The bitter lesson of misuse detection Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.563105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:38.761731Z digest=sha256:adf6d7200fb931728d2e93fbde1fb00caea0fd7d4ec6daa6c5cc3fee5499b421

Observation f8c82631-8dd1-4ee7-8b64-f927f3aa1a4a · outbound

This paper cites NeurIPS 2024 Datasets and Benchmarks Track, 2024.

The bitter lesson of misuse detection NeurIPS 2024 Datasets and Benchmarks Track, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.272682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:38.810817Z digest=sha256:78187927fa139f86bff0ad0099f4c6484467fbfba24a818fc9c1840170f05a8f

Observation 039fb256-65d4-4e2c-9e55-e0bbc9b68aee · outbound

This paper cites "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models.

The bitter lesson of misuse detection "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:38.885620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:38.885620Z digest=sha256:f80e2d7e75f775066a20fc2d624ca99ee026c58aba754b38f61554dd81dc296a

Observation 52615106-1ebe-4879-9046-de714e9e8211 · outbound

This paper cites SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal.

The bitter lesson of misuse detection SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.027078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.027078Z digest=sha256:585fdbdd3a47c65812bcecff5c7e7f2110ee2a245ac5458c6028404b620606ce

Observation df1cfef7-9e07-4e97-95c8-dabb2346ac2e · outbound

This paper cites AdvBench: Universal and Transferable Ad- versarial Attacks on Aligned Language Models,.

The bitter lesson of misuse detection AdvBench: Universal and Transferable Ad- versarial Attacks on Aligned Language Models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.058381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:39.230798Z digest=sha256:288d1d53b1b79e3ee1ccb010069fb03294b0cd64961ca1dbab1b532354289da8

Observation 23662cfe-a460-48e7-96b9-ee0dac2f98c2 · outbound

This paper cites CatQA: A Dataset for Categorizing Questions as Safe or Unsafe,.

The bitter lesson of misuse detection CatQA: A Dataset for Categorizing Questions as Safe or Unsafe,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.891885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:39.311873Z digest=sha256:14d2590b8a03f207f39febdb4581bc8ad9d2b0df973c34a8678fd40cdd3ea2ce

Observation f1c918c9-da12-4d60-8c4e-c6abb5eddbd7 · outbound

This paper cites Do Not Answer: Testing AI Refusal to Unsafe Questions,.

The bitter lesson of misuse detection Do Not Answer: Testing AI Refusal to Unsafe Questions,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.701691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:39.365817Z digest=sha256:289dbcb5a14cc38fdfcf079ce19aea17645159f0107d32b846065c63b31b8a62

Observation c0105b1f-cfac-43dd-8207-35d7d779c5c9 · outbound

This paper cites HH-RLHF: Helpful and Harmless Reinforcement Learning from Hu- man Feedback,.

The bitter lesson of misuse detection HH-RLHF: Helpful and Harmless Reinforcement Learning from Hu- man Feedback,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.506895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:39.472725Z digest=sha256:edc638b1e45c46328d37995f26100245863d7664a8c9089ca46a22e909189364

Observation 8543b09b-adba-4383-be52-c9d1d4dff676 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

The bitter lesson of misuse detection HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.588048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.588048Z digest=sha256:4c8ea7f5ff8f962253b1c16dba3ab6cf24f99b36b682972059f5393c4ff5c580

Observation 9a3352f7-c722-4e0a-92c8-3a166a2de097 · outbound

This paper cites A StrongREJECT for Empty Jailbreaks.

The bitter lesson of misuse detection A StrongREJECT for Empty Jailbreaks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.703539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.703539Z digest=sha256:d332055fd973079b548d796f65359835d8851c448cda7ce97d1a9d97d6f64a7c

Observation f6b9e1a3-e909-4b4a-9aab-9f0e404c9aac · outbound

This paper cites The Twelfth International Conference on Learning Representations, 2024.

The bitter lesson of misuse detection The Twelfth International Conference on Learning Representations, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.239110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:39.798982Z digest=sha256:c1dbc221843114cfc6cd018cf6e53f59bfa166a97db008d470c64bf3ae4982d6

Observation 82e96a74-af9e-44c2-99d9-b2a6927de0eb · outbound

This paper cites an unresolved cited work.

The bitter lesson of misuse detection Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:18:42.037851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:39.887601Z digest=sha256:dc7c1e727c7e608797d24550add11f316384556a6cc064dea4d977a6740b7ce6

Observation 219f32df-033c-482f-bf49-dc911055af8f · outbound

This paper cites an unresolved cited work.

The bitter lesson of misuse detection Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:18:41.743137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:40.022590Z digest=sha256:f3acf3f7556b3e1c6cc9195c0c586b61c18720269125f1ad71e420f13d8a7ef0

Observation 7efef56f-9ccc-4f3e-b0c4-589bac84fd58 · outbound

This paper cites DeepInception: Hypnotize Large Language Model to Be Jailbreaker.

The bitter lesson of misuse detection DeepInception: Hypnotize Large Language Model to Be Jailbreaker

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.154024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.154024Z digest=sha256:b334c92e599719138fd6b4f30d383a670d7eddf95cf12eb1ace10c31dbaee2d0

Observation 05f8fcb2-eb3f-441e-99f1-8a526aa9d2fe · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

The bitter lesson of misuse detection Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.283148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.283148Z digest=sha256:5cb65ec297b262eb7e748d68b2c6f6adf1184830b9b70247152c2c4ac13dfbc5

Observation 25f66079-b23c-41d8-b039-4b95459d3b93 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

The bitter lesson of misuse detection Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.429244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.429244Z digest=sha256:c3d56d0f56e45bae34eab5c3b3cc0bdf217b2d853c04304ea3dfdd31e3f84e2e

Observation cdba531b-0cde-4c13-b6f1-c526f77578c9 · outbound

This paper cites AI Control: Improving Safety Despite Intentional Subversion.

The bitter lesson of misuse detection AI Control: Improving Safety Despite Intentional Subversion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.580182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.580182Z digest=sha256:37d9251561a461dc41d079ef7fe3a4ab1bd4e5c38e43383d94293ebc7dfc2de0

Observation a53b041b-1dfd-4607-9626-78256768fe15 · outbound

This paper cites Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?.

The bitter lesson of misuse detection Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.663885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.663885Z digest=sha256:a26a09c8b5b7098c613e8748f4861db2a0d51607231a8da937aeb515a1ef3088

Observation 15b373f1-d82b-4fc5-b86b-a2de051580f3 · outbound

This paper cites Is this prompt harmful or not?.

The bitter lesson of misuse detection Is this prompt harmful or not?

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:41.442675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:18:40.850181Z digest=sha256:616389e5bcda2982cb4d9ca1359b37d266d162ca2bdb62418df33d40d7ada4d3

Pith citing papers

No inbound Pith citation observations are available.