Pith. sign in

Paper Citation Record · LEDGER

Challenges in Detoxifying Language Models

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2109.07445.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2109.07445 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:20:46.811816Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T13:33:28.439743Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bc9dd14a-fc78-4273-92cd-32bd42afb392 · inbound

A General Language Assistant as a Laboratory for Alignment cites this paper.

A General Language Assistant as a Laboratory for Alignment Challenges in Detoxifying Language Models

Reference 252

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:22:59.343502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-11T14:22:57.925354Z digest=sha256:2d589897a75f35730a4d227b7c6e7157cd217e632bc8d29cc195ca18b3a24890

Observation 92ed9766-bda1-4a38-9148-66d1467b19b2 · inbound

Ethical and social risks of harm from Language Models cites this paper.

Ethical and social risks of harm from Language Models Challenges in Detoxifying Language Models

Reference 290

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:24:30.322596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-11T18:24:28.835688Z digest=sha256:764873143082dd1515bb82ce659db88fffa34fd08b9ffa52b9dcce65c9cf1a96

Observation 186a5027-f14d-4885-85f4-8e785660a588 · inbound

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model cites this paper.

Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model Challenges in Detoxifying Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:14:26.647929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-24T12:10:49.690618Z digest=sha256:daa9be798fa78c9164c218982f6d67a2877b1317fa59f145bd59bdf029eb2b2f

Observation ee1b8db8-7624-4e50-87db-a35cc5a4277a · inbound

GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts cites this paper.

GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts Challenges in Detoxifying Language Models

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:25:21.123453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T06:25:20.966510Z digest=sha256:a21265c77ab6d4967bffb7d8317edd0fc691d8b5555dbaad4aee795cae8d2fdd

Observation fa4c0692-db09-4855-b1fc-85ec18622ced · inbound

TrustLLM: Trustworthiness in Large Language Models cites this paper.

TrustLLM: Trustworthiness in Large Language Models Challenges in Detoxifying Language Models

Reference 246

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.696987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:d79143709f5d839c0b7cbd06c1ce619a56dfaedac21ba81f8f93cf8eda436549

Observation 58f2a9d9-b396-4cef-a4ac-d270f48282ee · inbound

Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models cites this paper.

Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models Challenges in Detoxifying Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-12T05:33:18.447230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:33:18.447230Z digest=sha256:5a3bf3dcdd52938cce9b89954e51b71fec1da75424952bb351ee54d9a5b7a471

Observation b1fb9307-9bd1-4ebe-b733-2d6a47cdc529 · inbound

The Generative AI Ethics Playbook cites this paper.

The Generative AI Ethics Playbook Challenges in Detoxifying Language Models

Reference 126

Resolution
unresolved
no resolver link, observed 2026-08-11T13:15:10.934737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:15:10.934737Z digest=sha256:a2c4b2549ddac98ba42497f7451a02b0b013f8a534a0b15ccb0deb1c4c1cf729

Observation ee32c18a-1f36-4b04-b4aa-bc5d59673f41 · inbound

HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns cites this paper.

HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Challenges in Detoxifying Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-10T11:03:29.226885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:03:29.226885Z digest=sha256:acaca7ae37f9562926b429f4b9528097c359307208a7512cf5637268c7c3da38

Observation 3a2802d4-df00-42b2-a5dd-6af5299bd77b · inbound

ThreMoLIA: Threat Modeling of Large Language Model-Integrated Applications cites this paper.

ThreMoLIA: Threat Modeling of Large Language Model-Integrated Applications Challenges in Detoxifying Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T10:20:46.811816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:20:46.811816Z digest=sha256:24bf4051b1e7e863c86799fd27081b6156291f04ccff4f2fc265ce5689bb8b31

Observation 76698800-6522-4e0e-aaad-01e1f260eb78 · inbound

Moderating Harm: Benchmarking Large Language Models for Cyberbullying Detection in YouTube Comments cites this paper.

Moderating Harm: Benchmarking Large Language Models for Cyberbullying Detection in YouTube Comments Challenges in Detoxifying Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:05.971383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:05.971383Z digest=sha256:4faa04a7588aa8da823a23987bcccea87acb427214fc757d053a8e9df8863182

Observation d6f11497-6607-47b9-81c4-bce4f86b16c5 · inbound

When Large Language Models Meet Law: Dual-Lens Taxonomy, Technical Advances, and Ethical Governance cites this paper.

When Large Language Models Meet Law: Dual-Lens Taxonomy, Technical Advances, and Ethical Governance Challenges in Detoxifying Language Models

Reference 194

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:11.264262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:11.264262Z digest=sha256:4aaadfc92048a614581880885b67d31073065b1dcf979aa4a62c4e7b20514043

Observation 1d8c8e90-6eb3-4e3b-8b1a-518f1a939266 · inbound

Customize Multi-modal RAI Guardrails with Precedent-based predictions cites this paper.

Customize Multi-modal RAI Guardrails with Precedent-based predictions Challenges in Detoxifying Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T17:47:15.398133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:47:15.398133Z digest=sha256:6e9d042b98d63afd6d034e7606bb66da3bb1578f471649df262a1236b5e29450

Observation eabeb2e8-80d5-4dc2-b9c7-1693c012adc9 · inbound

Model Misalignment and Language Change: Traces of AI-Associated Language in Unscripted Spoken English cites this paper.

Model Misalignment and Language Change: Traces of AI-Associated Language in Unscripted Spoken English Challenges in Detoxifying Language Models

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T10:23:58.094942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:23:58.094942Z digest=sha256:5605aad75c568dee74da834f0bf7e33fc40d07885fc27174f89d2d7bed9309d1

Observation 3e9ac3c4-409c-4884-a32f-97a52c9211df · inbound

Ethics Testing: Proactive Identification of Generative AI System Harms cites this paper.

Ethics Testing: Proactive Identification of Generative AI System Harms Challenges in Detoxifying Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:56:08.093038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-09T20:49:22.147548Z digest=sha256:707ef2e43fc9091c26ba7d18f05cadaf750e31b1403daaf5aac799f8f01171c3

Observation 12ba0c4b-f95f-4d5a-8c36-19e1d2ec3a9b · inbound

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models cites this paper.

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models Challenges in Detoxifying Language Models

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.441271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-29T13:25:40.986336Z digest=sha256:4645dc831f486cc96bca81246f5708529464b2c1df5e6ae31f201c909c030a0a