Pith. sign in

Paper Citation Record · LEDGER

MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2406.17806.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.17806 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T19:45:20.041988Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T07:29:39.438493Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dc1d4cc0-e2f8-4ca9-ab34-28291416557f · inbound

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations cites this paper.

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 160

Resolution
unresolved
no resolver link, observed 2026-08-07T19:45:20.041988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:45:20.041988Z digest=sha256:e4044b096499e68fa51fd441503c34faada259b47b719a239446dcda163fa8ce

Observation 2335c84e-1709-42ae-983a-0070b3ad436f · inbound

VSCBench: Bridging the Gap in Vision-Language Model Safety Calibration cites this paper.

VSCBench: Bridging the Gap in Vision-Language Model Safety Calibration MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:12.324886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:13:12.324886Z digest=sha256:a7fb63409416deb46c9aea8230515cb476d8fe936dbdb951db9cf571a7f002d7

Observation b48dfcef-a75e-4f09-bc53-c11b55f00d6b · inbound

USB: A Comprehensive and Unified Safety Evaluation Benchmark for Multimodal Large Language Models cites this paper.

USB: A Comprehensive and Unified Safety Evaluation Benchmark for Multimodal Large Language Models MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:14.833561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:15:14.833561Z digest=sha256:5b537f3e8ffdea544bce29a176fc58f2b8aa8b790e69ae45e5f9a085db44869a

Observation 5bf4058d-6ff7-4787-ba6b-0bf388df51a8 · inbound

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal cites this paper.

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T23:31:54.525017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T23:27:52.438709Z digest=sha256:48e068bd54d0f2b775e9a699016efb556e5bcc3d30a933f2fac83994dfb0c4d0

Observation a20cbdb3-039b-45da-8e39-2b743bc0d66b · inbound

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs cites this paper.

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:35:57.449993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:04:05.157103Z digest=sha256:c93edfca192afc2bf9c1333454b85c04cadf0e726950c76fe4c2548321ed761e

Observation 750b19ee-79dd-42d9-8686-5305490996d1 · inbound

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models cites this paper.

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:57:27.607216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T06:56:10.053418Z digest=sha256:80d6299346d0a52dff19fc0a998c27e027d01639605c96817a926870bc98e03f

Observation 0a994371-8262-4f02-857a-0bbed69c171c · inbound

Auditing Multimodal LLM Raters: Central Tendency Bias in Clinical Ordinal Scoring cites this paper.

Auditing Multimodal LLM Raters: Central Tendency Bias in Clinical Ordinal Scoring MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:39:09.819823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:38:17.939704Z digest=sha256:5b0b9106cca5c55cf5fcfa87497731c9ea29980779fdb6f24cdfa3df99aa9fd9

Observation 3a67d623-c920-4217-9c14-e51f9e3067fb · inbound

AOR-Bench: Do Large Audio Language Models Over-Refuse Pseudo-Harmful Queries? cites this paper.

AOR-Bench: Do Large Audio Language Models Over-Refuse Pseudo-Harmful Queries? MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:29:39.439889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T13:22:12.541923Z digest=sha256:6dcadde8e038017ca197eb4a997426273a42aa1536a2fa3a412b9eac2a6f0c0c

Observation 61c847a2-bb09-4873-9bf1-2e5d82d57f48 · inbound

Securing Multimodal AI through Internal Information Decomposition cites this paper.

Securing Multimodal AI through Internal Information Decomposition MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T15:03:00.136710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:03:00.136710Z digest=sha256:9f807fb277e4380207ff193a498635c1055c89adfc989c606d866adc76583591

Observation 2bd241c8-8b8f-413a-822c-87815f35815d · inbound

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG cites this paper.

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-02T10:20:56.079722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:20:56.079722Z digest=sha256:d79a4447d15460e04a2c704e183e72a7e312698654d9a80ce266eb2a1126b13a