Pith. sign in

Paper Citation Record · LEDGER

PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2504.04377.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.04377 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:25:21.963287Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T10:04:51.407053Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b709662c-369e-4fde-8253-b4c71046e81e · inbound

The Problem with Safety Classification is not just the Models cites this paper.

The Problem with Safety Classification is not just the Models PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T12:25:21.963287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:25:21.963287Z digest=sha256:03f82ba3e876dfacc28e4b36709a13269b36adcca921f076882f8cf264a42c2c

Observation c8c6af8d-90d6-4fe4-9965-9e064e020637 · inbound

Tears or Cheers? Benchmarking LLMs via Culturally Elicited Distinct Affective Responses cites this paper.

Tears or Cheers? Benchmarking LLMs via Culturally Elicited Distinct Affective Responses PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T09:43:43.699290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:43:43.699290Z digest=sha256:652dd640975e2ef924715665d8687c9e96ca7318aa1a9a4e1d8e9821d2355685

Observation 1d797595-cba1-4718-a510-d988582b7e40 · inbound

YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models cites this paper.

YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T08:52:34.439908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:52:34.439908Z digest=sha256:343305597f6949afbb73ac729d270f720e331d0fe8d2e0f9da2db1149ffd8e1c

Observation 6c121afe-6766-490b-a0f9-3af68536ad04 · inbound

GuardReasoner-Omni: A Reasoning-based Multi-modal Guardrail for Text, Image, Video, and Audio cites this paper.

GuardReasoner-Omni: A Reasoning-based Multi-modal Guardrail for Text, Image, Video, and Audio PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T05:07:09.148505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:07:09.148505Z digest=sha256:2c2e89bbfb508dec619f50b718295aa7dd94f0b318be12d531691683fe0e0acb

Observation 98b784a4-5403-4266-a7de-6c23d804a7f0 · inbound

Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming cites this paper.

Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T11:43:31.089245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T11:43:31.089245Z digest=sha256:ca1bff474fdc15967a58f93f69b522c032e7527dfc1fcec3766c28a59d9fb729

Observation 6f7cd010-d28e-4448-bb35-d9bb0b7f5ff3 · inbound

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts cites this paper.

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:08:25.535585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T09:07:57.713675Z digest=sha256:95a43ceb58e7f5f5839e9fdf7a627753baffd30f411f44d3f7e902868eaf6150

Observation 1e9acd6d-91b0-4870-a830-94de92d2b6cd · inbound

LLM Safety From Within: Detecting Harmful Content with Internal Representations cites this paper.

LLM Safety From Within: Detecting Harmful Content with Internal Representations PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 74

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:20:23.043286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T04:33:54.058475Z digest=sha256:b0ee47deabb201887631512243971d01ef62bd3f04aff78ef87efb17e4ab7ece

Observation 965eb38b-1449-4aa6-a3f3-6a550707c1c6 · inbound

GLiNER Guard: Unified Encoder Family for Production LLM Safety and Privacy cites this paper.

GLiNER Guard: Unified Encoder Family for Production LLM Safety and Privacy PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:01:05.758750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T16:49:28.684979Z digest=sha256:1687dc69fa9815942f434be33bcf559cc54678e3a9f6789fe3ba58d4650c176a

Observation 5976ac45-e50a-40ca-8468-eb936b0c95db · inbound

GLiGuard: Schema-Conditioned Classification for LLM Safeguard cites this paper.

GLiGuard: Schema-Conditioned Classification for LLM Safeguard PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:20:55.245831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-11T03:19:11.495230Z digest=sha256:1b26ea0446907f4d45c80f81dfbfea64e479402a651ed72488e7a3c84b51c3d1

Observation 054ecdd2-5c5d-46c4-9cc7-7eadc66fd1f3 · inbound

LPG: Balancing Efficiency and Policy Reasoning in Latent Policy Guardrails cites this paper.

LPG: Balancing Efficiency and Policy Reasoning in Latent Policy Guardrails PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-19T23:43:17.946343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T23:42:58.135553Z digest=sha256:e67b64571172ad6293a2820281e1e0552fe71af0a0613a4bdd5dd35b2e0ccf05

Observation b07cbb8c-ce60-45b4-b5f7-947e2b5a6fad · inbound

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models cites this paper.

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.462067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-29T13:25:40.986336Z digest=sha256:8b91b70be26c47c9ff2a22fc2f630885ec86efdbc98cd05cc2ece0ae5ceea83f

Observation e9dd7f5c-8310-46be-815c-be9f007e2671 · inbound

Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content cites this paper.

Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-29T09:13:15.985122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T09:11:58.843585Z digest=sha256:003d1dd75308cf4f6f4c9c68745a482c8ba6efe9830cea4d650d2347d6688439

Observation 8fbdfea1-a8d5-4a05-bc78-6636fffc79f8 · inbound

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs cites this paper.

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:16:34.771753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T09:21:57.373862Z digest=sha256:ebdc2658361fdbbc2f5ced5f4f2929059b1f8f95a323f06b27836d56bcf79587

Observation 836768b1-b8bc-495c-8f36-13c04c510e60 · inbound

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models cites this paper.

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:30:07.670843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-25T21:16:36.392606Z digest=sha256:0378dccb3b9226aef1e15d30eb2460d1b2de7d340d1df6aaa3d5b68d6ec6d18b

Observation 53e91dc3-6849-4756-93a7-79d1224a4297 · inbound

HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety cites this paper.

HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:48:32.674518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T14:38:55.045628Z digest=sha256:1223600a73347eb8f598110a58f78a648f56faf6869c6e8005ad2a2d4984224c

Observation c2c96ca1-9bcf-44a2-bfae-294d9b3a95d3 · inbound

DT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety Guardrail cites this paper.

DT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety Guardrail PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-08T10:04:51.410268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-08T10:00:02.677030Z digest=sha256:0814d2e2ee6907cee39c86f62b4860fbff98885854bdb6cac6939740598929d6

Observation 43e4a66e-4b89-4a16-8af5-b713bbfd82dc · inbound

RAIL Guard: Closing the Evaluation-to-Remediation Gap in Responsible AI for LLM Agents cites this paper.

RAIL Guard: Closing the Evaluation-to-Remediation Gap in Responsible AI for LLM Agents PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T12:50:38.561845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:50:38.561845Z digest=sha256:a43de56dc39b8dcba03493143e8e27a72a1f05c93b73c71652effa4bc500b1b7

Observation 3edaeb5b-24e8-4053-a614-0326e80b7c96 · inbound

JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety cites this paper.

JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T11:24:49.023248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T11:24:49.023248Z digest=sha256:ef24ea553f045681ce7efb32d0308c30a27a8c3ab07f9c15a36234d94750c27a