Pith. sign in

Paper Citation Record · LEDGER

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset

As of 14 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 1 inbound Pith citation observation for arXiv:2411.08243.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.08243 v3

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:52:34.686086Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T12:58:09.227648Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 478f6d08-7e46-4bfa-a518-916494590a2e · outbound

This paper cites Leveraging Large Language Models in Conversational Recommender Systems.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Leveraging Large Language Models in Conversational Recommender Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.637138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.637138Z digest=sha256:5a098f74e9875befb4b00938c28ac160708444f188064d735cd71bb2213d905d

Observation 7f6dae83-1fd2-49f2-ad7d-acbcc3d9751f · outbound

This paper cites BERTopic: Neural topic modeling with a class-based TF-IDF procedure.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset BERTopic: Neural topic modeling with a class-based TF-IDF procedure

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.646442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.646442Z digest=sha256:21414ea0db106d7fdc95ed1933b397f24bf4405d670e23e99e9a8bf705f5fd2b

Observation 9b0fee32-03aa-4a0f-9155-e1961ec42e11 · outbound

This paper cites The Curious Case of Neural Text Degeneration.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset The Curious Case of Neural Text Degeneration

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.651073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.651073Z digest=sha256:0b7c11e6ad3384d9068490066729e6aa18d0b54a325aade93ef2654ec8670cd4

Observation 2d055cf9-5346-4e77-be64-37b7d71acfff · outbound

This paper cites The Alignment Problem from a Deep Learning Perspective.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset The Alignment Problem from a Deep Learning Perspective

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.659958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.659958Z digest=sha256:e12af0e11fc831b8e5b0116968d7eefef5403781e0a5761fa84e7d30b3a743f7

Observation 841dca4e-2f1c-47a2-a249-54ac9aa5b15e · outbound

This paper cites GPT-4 Technical Report.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset GPT-4 Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.664347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.664347Z digest=sha256:b1fbdb6f488efd5a9c76e9d563fac9a4c5b2fe6507e12832d469cf7d4c139977

Observation 4f01ec54-941c-4e56-bd85-34519288d3cb · outbound

This paper cites Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.668288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.668288Z digest=sha256:96b518bb09c4b9af95b1bb1de83e608ba63c6d03e3ac01e5f48393375aef17cd

Observation 08d542fd-f972-452b-b354-6fab2bc2b4d1 · outbound

This paper cites Character-LLM: A Trainable Agent for Role-Playing.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Character-LLM: A Trainable Agent for Role-Playing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.672381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.672381Z digest=sha256:994672dd5210ba2bdc32c29a50e94bd014c9052338267de4155985455f7b2994

Observation 67b01a54-9a65-4e89-bea7-243e979e3c72 · outbound

This paper cites Fundamental Limitations of Alignment in Large Language Models.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Fundamental Limitations of Alignment in Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.676905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.676905Z digest=sha256:3cac2fbe9aaf4b2d50cbe4435b0d07f43f95ff3515e2ac48b5ee747d59bb87d8

Observation 7d3b536a-e6c6-4b8f-a88e-f2be84c5ca21 · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.681096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.681096Z digest=sha256:bbfd4f417f51e4b8635a75652a8398347a29f1ce5d4509b9d6da0e88fcee4444

Observation c7deb38d-0573-4119-9cc1-16579962f757 · outbound

This paper cites In total, the dataset contains 22K toxic prompts.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset In total, the dataset contains 22K toxic prompts

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:52:34.850548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T21:52:34.686086Z digest=sha256:94b3ba2d17893dd0a4b85817514d146f4cdff76fe8426d60bf3e168ee5ebe03f

Observation 02e1a373-9aea-4624-abf1-a42063183e9d · outbound

This paper cites Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.655368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.655368Z digest=sha256:b16707fbecd68a1fba55aea5dee1db4cc83c14bb963dc24b60311e8ac8d76156

Observation de241576-573b-43f5-962c-f0f635f22ce7 · outbound

This paper cites Safe RLHF: Safe Reinforcement Learning from Human Feedback.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Safe RLHF: Safe Reinforcement Learning from Human Feedback

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.632566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.632566Z digest=sha256:0395d3d657ace535cf65c54f7fc0e2f077fc1121e690b8ba0d6a46a138a9f0e4

Observation 801ac7b6-4366-48b5-88ff-447686eae7da · outbound

This paper cites CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.642407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.642407Z digest=sha256:6dde4b0db1318a441623c670c8825ed11f883e78ed332163a87b5d93d5a9a9ba

Pith citing papers

Observation c26fdc0a-8967-4c72-9735-541cb32951c2 · inbound

Discriminatory Compliance: How LLMs Answer Queries from Protected Groups cites this paper.

Discriminatory Compliance: How LLMs Answer Queries from Protected Groups Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-06-26T12:59:29.240410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-26T12:58:09.227648Z digest=sha256:6428d0700e07f88c742006ee65ee74ae24de2105a715fa5177fbeb20fd433950