Pith. sign in

Paper Citation Record · LEDGER

A Holistic Approach to Undesired Content Detection in the Real World

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2208.03274.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2208.03274 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:28:43.934564Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T13:06:58.683371Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2bba94ca-935b-4b55-abb2-43ad148bad9f · inbound

Ignore Previous Prompt: Attack Techniques For Language Models cites this paper.

Ignore Previous Prompt: Attack Techniques For Language Models A Holistic Approach to Undesired Content Detection in the Real World

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:59:31.495804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T13:59:31.213830Z digest=sha256:6d1a3d4cbff9cba8179f8b9534dd61c7f206353ec17ebbed341f687a8a7d7708

Observation 7571bf72-6d0d-4cc0-8123-3656d96f888e · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models A Holistic Approach to Undesired Content Detection in the Real World

Reference 128

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T06:38:36.858913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:a1f9d7926cb6bfd69a6b9421d4ecd7da11d014641114127e86cbccb34f1c4ad9

Observation fc3b47ae-d489-4c17-a7db-e115376c280c · inbound

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction cites this paper.

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction A Holistic Approach to Undesired Content Detection in the Real World

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:37:15.936802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T11:34:09.428653Z digest=sha256:165b54cb802b14344974aa959f4bd38ac1e650257e1a5422309e59b07431a958

Observation 95f93033-9560-45a8-bcae-3076a86fa639 · inbound

SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems cites this paper.

SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems A Holistic Approach to Undesired Content Detection in the Real World

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:43.934564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:28:43.934564Z digest=sha256:39c73791e740c0e7bbd69f4482bb3f5277a43342062c40cc87b30e2c9b44c6c8

Observation 17a36167-b357-4480-bee9-490901646137 · inbound

Guard Vector: Beyond English LLM Guardrails with Task-Vector Composition and Streaming-Aware Prefix SFT cites this paper.

Guard Vector: Beyond English LLM Guardrails with Task-Vector Composition and Streaming-Aware Prefix SFT A Holistic Approach to Undesired Content Detection in the Real World

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T14:50:30.992141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:50:30.992141Z digest=sha256:1183895125d033ffbab81077256777ac8d71011e2a3f86bd119061f216a01399

Observation 35db1525-7155-47e0-9ca4-3ac1bf0c7fa9 · inbound

GuardReasoner-Omni: A Reasoning-based Multi-modal Guardrail for Text, Image, Video, and Audio cites this paper.

GuardReasoner-Omni: A Reasoning-based Multi-modal Guardrail for Text, Image, Video, and Audio A Holistic Approach to Undesired Content Detection in the Real World

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T05:07:09.386730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:07:09.386730Z digest=sha256:03555bb586e9d4205647823989d5112938ae34251c77e0123b49074972f1f67b

Observation dc21f6c3-6724-43c2-8c5f-32cef1d4e304 · inbound

Response-Based Knowledge Distillation for Multilingual Jailbreak Prevention Unwittingly Compromises Safety cites this paper.

Response-Based Knowledge Distillation for Multilingual Jailbreak Prevention Unwittingly Compromises Safety A Holistic Approach to Undesired Content Detection in the Real World

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:28:48.799684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T01:27:16.967080Z digest=sha256:249e270376b9ae6febabe6333c6456c11216d146ebddc753330e1d071699f4e8

Observation 98b44a5c-3ee8-4942-9058-04915b276034 · inbound

Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming cites this paper.

Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming A Holistic Approach to Undesired Content Detection in the Real World

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T11:43:31.089245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T11:43:31.089245Z digest=sha256:6df4fe5a0746b4caa0864b69d39ec78eed6c0826fd8a1fade638dd23d2f447d2

Observation 945b456b-7538-433d-b7c0-7265147d2d27 · inbound

VoxSafeBench: Not Just What Is Said, but Who, How, and Where cites this paper.

VoxSafeBench: Not Just What Is Said, but Who, How, and Where A Holistic Approach to Undesired Content Detection in the Real World

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:24:22.129507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T10:19:28.041282Z digest=sha256:1c7c82a06e305ef36083c07dff3fdd43a1557221b27b4d777cedd729eedf0348

Observation d36ec0b3-12a0-4029-a2a9-def6c38c666e · inbound

Test-Time Safety Alignment cites this paper.

Test-Time Safety Alignment A Holistic Approach to Undesired Content Detection in the Real World

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:51:43.314183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T16:06:37.244288Z digest=sha256:1ad287b34b87890e406d2785c907d1cc5095e94969fd6d698c2ae0035dc71e0f

Observation b8c0d322-4c2c-4b4a-9de7-719b014adb7b · inbound

AI Content Moderation in Therapy Conversations cites this paper.

AI Content Moderation in Therapy Conversations A Holistic Approach to Undesired Content Detection in the Real World

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:03:58.416947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-29T21:02:26.865075Z digest=sha256:527e6d70e44712a297856d2c76e917420409bdc89d1f6396597e0c7296d5fe77

Observation 05c10576-7de2-48e1-9a74-96b54911bc59 · inbound

$D^2$-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing cites this paper.

$D^2$-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing A Holistic Approach to Undesired Content Detection in the Real World

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:03:59.808267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T22:03:25.762772Z digest=sha256:491ebc731aafeb5dcaae894792e917e3326c8a2f41ce3358a827e68cbfb09b5f

Observation 95c790ca-7fc2-4ab8-9e9c-0e341194407e · inbound

Learning from Mistakes: Can LLM Self-Recover after Misalignment? cites this paper.

Learning from Mistakes: Can LLM Self-Recover after Misalignment? A Holistic Approach to Undesired Content Detection in the Real World

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-13T18:51:10.298187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:51:10.298187Z digest=sha256:c24a744c2958e616b53913252c26656381dae19e7709880566b2536d291b0127

Observation a5eec1c6-b462-4521-b7db-96e4f953b050 · inbound

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety cites this paper.

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety A Holistic Approach to Undesired Content Detection in the Real World

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:52:55.802055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T00:46:03.210076Z digest=sha256:fda9c628d9f11bed7613e2170b77b5411ff3da2e8b90236a5f7f3ed2066aa742

Observation 4a973d41-9d02-47d2-a9bf-a133257b3289 · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models A Holistic Approach to Undesired Content Detection in the Real World

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:35:51.371484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:e660fc566cbe2bccb9a128791c7ce81f18c209d020d3ff2af1d7a51f92fe598e

Observation 425d607f-c7d4-44b5-9195-0eaf66fa9f7d · inbound

AI Native Games: A Survey and Roadmap cites this paper.

AI Native Games: A Survey and Roadmap A Holistic Approach to Undesired Content Detection in the Real World

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:58.684741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T12:59:20.908659Z digest=sha256:48bc5fa0b69f064442f05c0b71195024b59b0a3103a27defa2314e085ce084ea

Observation f3d89b71-7b73-46ad-abc0-ce6be5bdc526 · inbound

AI Native Games: A Survey and Roadmap cites this paper.

AI Native Games: A Survey and Roadmap A Holistic Approach to Undesired Content Detection in the Real World

Reference 89

Resolution
unresolved
no resolver link, observed 2026-07-12T09:30:07.729486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:30:07.729486Z digest=sha256:44eb148769ee4bca3ae3337351fea124372fe70d2a3180eed87522bca3e5cfa3