Pith. sign in

Paper Citation Record · LEDGER

MSTS: A Multimodal Safety Test Suite for Vision-Language Models

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2501.10057.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10057 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T08:23:30.999477Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T13:54:58.568078Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b221af96-e7ab-46a8-aa3d-05734cc10e1d · inbound

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation cites this paper.

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation MSTS: A Multimodal Safety Test Suite for Vision-Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:15:54.769687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T03:13:05.662351Z digest=sha256:6d71eb27b181a778b90b2c1e494669ebc9d0a16c7d98e1598ca9a1f756daa3eb

Observation 93084c44-2646-4e81-95f1-4b87d9d04c14 · inbound

No Safe Dose: How Training Data Drives Unsafe Image Generation cites this paper.

No Safe Dose: How Training Data Drives Unsafe Image Generation MSTS: A Multimodal Safety Test Suite for Vision-Language Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:43:28.536947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T13:41:53.768791Z digest=sha256:c6cf1574b66086ec7c7c1712f861f2953ff0ddd34bfb1701dc610556f808d64b

Observation cfb703dd-214c-4aeb-a752-1821d2412992 · inbound

RedVox: Safety and Fairness Gaps in Speech Models Across Languages cites this paper.

RedVox: Safety and Fairness Gaps in Speech Models Across Languages MSTS: A Multimodal Safety Test Suite for Vision-Language Models

Reference 147

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:59:52.818809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T04:37:00.399470Z digest=sha256:a0ec47fc954cbc2a41d42292a2b77e35997e5d053bc53707e1aeaba51494d17a

Observation a7068c13-9575-429a-9ccc-08bac30c12cc · inbound

Pluralis v0.1: Towards a Multicultural, Multimodal, Multilingual Benchmark for AI Risk and Reliability cites this paper.

Pluralis v0.1: Towards a Multicultural, Multimodal, Multilingual Benchmark for AI Risk and Reliability MSTS: A Multimodal Safety Test Suite for Vision-Language Models

Reference 125

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T13:54:58.569580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-08T13:50:54.083165Z digest=sha256:faebfce0c91ccfe2b2e40dee69eeaf572973dfbdfc124b78f3e989bc9744011b

Observation 2730f344-ed6c-46c2-9375-f2c98d0c3b49 · inbound

Multimodal Reward Hacking in Reinforcement Learning cites this paper.

Multimodal Reward Hacking in Reinforcement Learning MSTS: A Multimodal Safety Test Suite for Vision-Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T02:39:02.891861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T02:39:02.891861Z digest=sha256:bad2183541e532112f93af19f8fbf6819c5814c4edeb88abd191a79e56398dc8

Observation acff3c2e-36fe-41ef-85c3-3f6dad372087 · inbound

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure cites this paper.

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure MSTS: A Multimodal Safety Test Suite for Vision-Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T08:23:30.999477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:23:30.999477Z digest=sha256:78fffba1c2e08edb467d576f897b84b0925aefe2e71449bdbb907d81681f6e55