Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2310.17389.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:52:28.909375Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T14:48:32.654187Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 4ae480b6-9c61-470c-a90d-6399246e790d · inbound
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6049c2b-aeaa-44a8-8685-d434ac56dbb4 · inbound
ShieldGemma: Generative AI Content Moderation Based on Gemma ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e773cb32-b65a-4db0-8695-1cd4132b6e08 · inbound
SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise Use ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96d8ff20-356b-4be2-9aba-bb9b38955dea · inbound
Large Language Models Often Know When They Are Being Evaluated ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0f879a3-92f5-496f-abec-9b8f05638b4d · inbound
From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 899debdd-8dbc-4652-b50f-8fe1d3245bd9 · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a59b22d2-ab12-433d-8fca-fcf93f1968f3 · inbound
Attestable Audits: Verifiable AI Safety Benchmarks Using Trusted Execution Environments ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8945f4db-2a41-4c61-9a11-ef0ac1d66d3e · inbound
LLMs Encode Harmfulness and Refusal Separately ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d08457f-5a13-4da3-af0c-6a428b8406fa · inbound
From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe7c1cb3-98e9-4202-af1b-118f250474bf · inbound
GuardReasoner-Omni: A Reasoning-based Multi-modal Guardrail for Text, Image, Video, and Audio ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86848d71-d41c-4229-aecd-7ea9e6c58cbc · inbound
Less is Enough: Synthesizing Diverse Data in LLM Feature Space with Sparse Autoencoders ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7e3cf6a-63e5-4fe4-b5bc-3396ff7d194c · inbound
FedDetox: Robust Federated SLM Alignment via On-Device Data Sanitization ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14c65156-fe1a-4726-8f6c-d6fff637a2ff · inbound
Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f3e2d616-a3f3-4103-a614-aa060b0a2e05 · inbound
TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5b2a56b-da7c-4d7e-99ac-83fb41c9db39 · inbound
LLM Safety From Within: Detecting Harmful Content with Internal Representations ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8af5d2c-560f-4e6f-aa37-e9d5cb70633c · inbound
Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d758f06-8adc-4c14-8c4d-2f363f1f20ad · inbound
Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 01f579b9-361c-4712-a23f-9db28b5a9661 · inbound
Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 700c0b50-026d-4736-b3d9-31e4fa83de6f · inbound
When Youth Enter the Algorithmic Wild: Discovering and Understanding Potentially Harmful Teen Videos on Douyin and Kwai ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c61f411c-1406-4b77-8914-6f10225690b0 · inbound
Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 59558a22-481c-44f1-a762-506f6272a304 · inbound
HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 048e4edc-5e2d-46cd-aa71-38c0bc50b6ab · inbound
Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fd0eeb0-ba99-451f-a1ed-ebdbb1a87b24 · inbound
A Constitution-Grid Instrument for Data-Efficient RL Alignment (C-Guard) ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.