Pith. sign in

Paper Citation Record · LEDGER

ShieldGemma 2: Robust and Tractable Image Content Moderation

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2504.01081.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.01081 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:04:05.817517Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:09:54.641120Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 513fb847-4e87-4aea-995d-80b531f0503d · inbound

VLMs Can Aggregate Scattered Training Patches cites this paper.

VLMs Can Aggregate Scattered Training Patches ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:05.817517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:05.817517Z digest=sha256:256635d38def4a8518eb6460a62f2afcf4863ae8f5eda7e611c885331800e35d

Observation 565c71ff-9edc-481f-a63f-4b03b833470d · inbound

Personalized Constitutionally-Aligned Agentic Superego: Secure AI Behavior Aligned to Diverse Human Values cites this paper.

Personalized Constitutionally-Aligned Agentic Superego: Secure AI Behavior Aligned to Diverse Human Values ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:07.193249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:43:07.193249Z digest=sha256:7342a596f16b8b93cbffa931a4e14fe1caea2aaba2f9e8098206de509a294407

Observation 2a9a2bfd-6676-4cfc-a215-84e3404c3bd1 · inbound

Whose View of Safety? A Deep DIVE Dataset for Pluralistic Alignment of Text-to-Image Models cites this paper.

Whose View of Safety? A Deep DIVE Dataset for Pluralistic Alignment of Text-to-Image Models ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T17:08:20.214896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:08:20.214896Z digest=sha256:bdf204554a3a7a824b3be3c4ba4406a044684a9c1eab103ad1c3e74a449d751c

Observation 2600760a-ae98-4cd1-ba21-1b45ed2c279a · inbound

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models cites this paper.

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T12:42:36.804758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-18T12:41:48.620040Z digest=sha256:ab2fb7e5f61910696b53d0033d23ebec57eeb809af7e1f596b1a21bcde1c4802

Observation 67bafbe8-bdda-4c3d-b079-f68e55ebbcbc · inbound

SenBen: Sensitive Scene Graphs for Explainable Content Moderation cites this paper.

SenBen: Sensitive Scene Graphs for Explainable Content Moderation ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:01:01.787041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T16:51:01.201885Z digest=sha256:6fbfa9f51658d3f39f83e3482f77ba67103868bbdef5d8cc74ebbc54f237f93e

Observation 979ac981-a85c-492d-89d3-fdc54b545123 · inbound

SenBen: Sensitive Scene Graphs for Explainable Content Moderation cites this paper.

SenBen: Sensitive Scene Graphs for Explainable Content Moderation ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T23:42:32.347933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T23:42:32.347933Z digest=sha256:66648abf9b4dc009c3efe4e5b166ac5284426dc3166617d2466b3051bd2a3b5f

Observation 0c507a52-f587-4a0c-b449-40a5acdee1e3 · inbound

SafeLens: Deliberate and Efficient Video Guardrails with Fast-and-Slow Screening cites this paper.

SafeLens: Deliberate and Efficient Video Guardrails with Fast-and-Slow Screening ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:08:21.190297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T14:03:56.859481Z digest=sha256:3dbaca999ce742e523ab4e3d14fb9dade7d42786ab90e35e2db46f9344256983

Observation 39cc81f6-7e92-4b9c-bec5-2603223e608c · inbound

Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South cites this paper.

Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:08:07.042646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:06:57.555070Z digest=sha256:7dbe7aa6ce79677fb0d5a45f5fc423cdd008aebd6c78ff74b9fe8f35b77a0782

Observation e011571c-b4a4-4314-b45c-b3ef1731047a · inbound

Boundary-targeted Membership Inference Attacks on Safety Classifiers cites this paper.

Boundary-targeted Membership Inference Attacks on Safety Classifiers ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:11:17.541591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T08:06:45.896332Z digest=sha256:f5f4e963dabcb5bbaabb1df80e0a429a38d6987dca756c7680ee1483615a11c9

Observation b6096ed2-305c-4eb0-bcc7-4120c3746323 · inbound

Boundary-targeted Membership Inference Attacks on Safety Classifiers cites this paper.

Boundary-targeted Membership Inference Attacks on Safety Classifiers ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:24.307746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-25T05:42:55.999769Z digest=sha256:80fc645e8dcad5a3743adb60c177c3a3af17f1b15ac597819db3136a477a3d85

Observation d826f66f-91a9-43c1-bd93-c860745488c3 · inbound

$D^2$-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing cites this paper.

$D^2$-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T22:03:59.797888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T22:03:25.762772Z digest=sha256:1c7657d8a7dc7182d7bf0275de30f702d44671b66c8d29e41cdaff9b75781faf

Observation de449355-e559-4c24-be03-4f26a87095cd · inbound

No Safe Dose: How Training Data Drives Unsafe Image Generation cites this paper.

No Safe Dose: How Training Data Drives Unsafe Image Generation ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:43:28.537424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T13:41:53.768791Z digest=sha256:2587f26919537a9440173649b06b1fb48ee3fe27cc33745a468c3fd9e311718e

Observation 46013750-9652-46f7-955c-fef2a5b8773d · inbound

MIRAGE: Protecting against Malicious Image Editing via False Moderation cites this paper.

MIRAGE: Protecting against Malicious Image Editing via False Moderation ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:54.643230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-26T01:52:12.291420Z digest=sha256:d1b400cdb95fd79d92ecca0181d8e94ae9d7f46c4cf32685bcce5c180c6a8f72

Observation 373824c2-71b4-4cbe-b72d-ac420de6a596 · inbound

MIRAGE: Protecting against Malicious Image Editing via False Moderation cites this paper.

MIRAGE: Protecting against Malicious Image Editing via False Moderation ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:13:53.510290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T04:46:56.552601Z digest=sha256:30469685b8faf4516d574cc5905db994580c536b79ed2cf43c7d8b8b3352cbb4

Observation 6ba0bf05-8920-416b-95e3-1cf68579ff61 · inbound

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing cites this paper.

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T06:34:18.821395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T06:31:55.631719Z digest=sha256:6f0f8208106cc5671bb0844eab87bf09d8b57d41e99789bce2919bec842f2211

Observation 2c519ab3-a279-4d56-80a5-5161116b8721 · inbound

Every Sample Counts: Supervised Fine-Tuning of Language Models with Pointwise Constraints cites this paper.

Every Sample Counts: Supervised Fine-Tuning of Language Models with Pointwise Constraints ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 97

Resolution
unresolved
no resolver link, observed 2026-07-13T05:26:06.829382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:26:06.829382Z digest=sha256:f49613b7cc8619d56abf3d4b8f2fd2c6efb9bbf93f5abbfb0562378ab81dad09

Observation 37172ff4-17a5-4710-a802-65f27e690df0 · inbound

Old Tricks, New Models: How Simple Image Transformations Break Modern AI-based Content Moderation cites this paper.

Old Tricks, New Models: How Simple Image Transformations Break Modern AI-based Content Moderation ShieldGemma 2: Robust and Tractable Image Content Moderation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T15:15:01.267197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T15:15:01.267197Z digest=sha256:6865499e05ed33f093dd7eb11d4ed0baa15a0d617ab09a4d35e8d2966d9f3735