Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:52:34.686086Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 1 inbound Pith citation observation for arXiv:2411.08243.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:52:34.686086Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-26T12:58:09.227648Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
13 of 13 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 478f6d08-7e46-4bfa-a518-916494590a2e · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Leveraging Large Language Models in Conversational Recommender Systems
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f6dae83-1fd2-49f2-ad7d-acbcc3d9751f · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset BERTopic: Neural topic modeling with a class-based TF-IDF procedure
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b0fee32-03aa-4a0f-9155-e1961ec42e11 · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset The Curious Case of Neural Text Degeneration
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d055cf9-5346-4e77-be64-37b7d71acfff · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset The Alignment Problem from a Deep Learning Perspective
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 841dca4e-2f1c-47a2-a249-54ac9aa5b15e · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset GPT-4 Technical Report
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f01ec54-941c-4e56-bd85-34519288d3cb · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08d542fd-f972-452b-b354-6fab2bc2b4d1 · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Character-LLM: A Trainable Agent for Role-Playing
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67b01a54-9a65-4e89-bea7-243e979e3c72 · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Fundamental Limitations of Alignment in Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d3b536a-e6c6-4b8f-a88e-f2be84c5ca21 · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7deb38d-0573-4119-9cc1-16579962f757 · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset In total, the dataset contains 22K toxic prompts
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 02e1a373-9aea-4624-abf1-a42063183e9d · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de241576-573b-43f5-962c-f0f635f22ce7 · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Safe RLHF: Safe Reinforcement Learning from Human Feedback
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 801ac7b6-4366-48b5-88ff-447686eae7da · outbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c26fdc0a-8967-4c72-9735-541cb32951c2 · inbound
Discriminatory Compliance: How LLMs Answer Queries from Protected Groups Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.