Pith. sign in

Paper Citation Record · LEDGER

Semantic Robustness Certification for Vision-Language Models

As of 5 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2606.18839.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.18839 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-26T21:32:54.522220Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact8
  • verified fuzzy0
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d781d1b0-3d77-4ab3-933a-029a48e9133e · outbound

This paper cites GPT-4 Technical Report.

Semantic Robustness Certification for Vision-Language Models GPT-4 Technical Report

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T23:59:07.057706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:c4846c7c451e83cc5d11acdff9ffd1289ccdb63cc6b3c83acbf2d0c7c2cdb1bc

Observation 66a59692-bfd6-4c34-a769-68a91253990f · outbound

This paper cites Food-101– mining discriminative components with random forests.

Semantic Robustness Certification for Vision-Language Models Food-101– mining discriminative components with random forests

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-26T21:32:54.522220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:5ddfdc0de841a3218d3ebc8bf772d74a5d0b2e09a2002e841fb6f3b64b1194b9

Observation 362f6d85-831a-4e85-8032-4032cd63a291 · outbound

This paper cites Interpreting CLIP: Insights on the Robustness to ImageNet Distribution Shifts.

Semantic Robustness Certification for Vision-Language Models Interpreting CLIP: Insights on the Robustness to ImageNet Distribution Shifts

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:59:07.073171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:e2c279522f6c3529daba62a68b6bd4c9029ce597dc2c121296dd8d657b4d886d

Observation 52301229-1dc5-4d56-8335-65f646cba09d · outbound

This paper cites Complete Verification via Multi-Neuron Relaxation Guided Branch-and-Bound.

Semantic Robustness Certification for Vision-Language Models Complete Verification via Multi-Neuron Relaxation Guided Branch-and-Bound

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:59:07.081746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:9943d08a68241f4ac79aaff5009ad52111a92506918b9c1817a236cc739ce134

Observation 365625cd-f28f-4524-b722-37a82054181a · outbound

This paper cites Seed1.5-VL Technical Report.

Semantic Robustness Certification for Vision-Language Models Seed1.5-VL Technical Report

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:59:07.085282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:09964f3835fb0584a80dd85548da0f879cd0263983a843086d2c32cea49b48fb

Observation ba317e7a-49f1-4aa6-9164-6f12c495b091 · outbound

This paper cites Mmt-ard: Multimodal multi-teacher adversarial distillation for robust vision-language models.

Semantic Robustness Certification for Vision-Language Models Mmt-ard: Multimodal multi-teacher adversarial distillation for robust vision-language models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:59:07.078587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:3f833f3d38794acf8ed0d6120ce1595ff3095cce1243ad0c6fe0e817fdae3fe9

Observation 2bae19ac-3899-4aee-8788-6dba7a799244 · outbound

This paper cites Fine-Grained Visual Classification of Aircraft.

Semantic Robustness Certification for Vision-Language Models Fine-Grained Visual Classification of Aircraft

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:59:07.089292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:586b6a258c2cccd9e431e22af01ff47067aa24476f94933d24c9c824f2371eb2

Observation 89952cc8-2263-4c1c-b85d-c46a9375ba8c · outbound

This paper cites Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models.

Semantic Robustness Certification for Vision-Language Models Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:59:07.065828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:a5870bc2e8213f53cdb9bcfa9d97cc272b6f8f79454104ca86c3d048837d0718

Observation 3d5a0d0f-55a5-4b0b-935d-b57c087bef80 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Semantic Robustness Certification for Vision-Language Models UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:59:07.077457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:592e30506920f8b1d2a6da19865b6329985cdf57ee492decfeb1505b6118accd

Observation bf11003b-d52f-4c1e-aa4e-e495a2eadfb7 · outbound

This paper cites Local path inte- gration for attribution.

Semantic Robustness Certification for Vision-Language Models Local path inte- gration for attribution

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:59:07.073552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:a320fe4d4a23a75879f5e726b99f1a0020b362e2c36cc742843fe2a6618618e3

Observation fd3bbc3c-8a0b-4237-9c61-275a5a12034f · outbound

This paper cites Ants: Adaptive negative textual space shaping for ood detection via test-time mllm understanding and reasoning.

Semantic Robustness Certification for Vision-Language Models Ants: Adaptive negative textual space shaping for ood detection via test-time mllm understanding and reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:59:07.068467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:00e936f2450f3d64a91a9b5bb0a41fa4c35207d9ae2ee386e471c860cae30d26

Observation 89513444-b139-46c0-aedd-fb6ad45b2fcb · outbound

This paper cites an unresolved cited work.

Semantic Robustness Certification for Vision-Language Models Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-26T21:32:54.522220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:72705bc37501faeefc359f5ec45e52de7498db087a736754051d050db925f718

Observation 536ae4fb-f0c4-4203-b35a-926900fc574c · outbound

This paper cites We therefore use multimodal large language models (MLLM) (e.g., GPT models (Achiam et al.,.

Semantic Robustness Certification for Vision-Language Models We therefore use multimodal large language models (MLLM) (e.g., GPT models (Achiam et al.,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-26T21:32:54.522220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:c513dcbf6db33ce8af43acce9a46e4ea10366b5df74680d0d7b375bb277bf009

Observation 53260d3c-83ef-4280-8c3b-4e42e916b035 · outbound

This paper cites a photo of a [attribute] [class].

Semantic Robustness Certification for Vision-Language Models a photo of a [attribute] [class]

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-26T21:32:54.522220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T21:32:54.522220Z digest=sha256:a51862f3ef82a878a3db366e0eb34278b72c790e4ad4a7e8c6a3228c6fd7aec2

Pith citing papers

No inbound Pith citation observations are available.