Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T16:37:19.310821Z
Paper Citation Record · LEDGER
As of 3 August 2026, this Paper Citation Record lists 6 of 6 outbound references and 1 inbound Pith citation observation for arXiv:2604.12371.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T16:37:19.310821Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T02:44:13.401198Z
A source-named dated measurement, never combined with another source.
Source: cited_works
6 of 6 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d4ea2ad9-f244-488b-a5b1-93850ff9909e · outbound
Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models Jina CLIP: Your CLIP Model Is Also Your Text Retriever
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 24cfb540-00e6-46f3-afe0-87e394e51efb · outbound
Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models Salad-bench: A hierarchical and comprehensive safety benchmark for large language mod- els
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 10b826a8-6271-408b-9c33-79d68de62f2f · outbound
Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models VLM-Guard: Safeguarding Vision-Language Models via Fulfilling Safety Alignment Gap
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 421ca7c1-2f1d-42c6-8f68-cfb678f97af7 · outbound
Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models Typographic attacks in a multi-image setting
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0978cd53-98d4-4bb0-ba84-f61af5960e47 · outbound
Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models SCAM: A real-world typographic robustness evaluation for multimodal foundation models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 985309d2-d4a7-4328-9fdb-93f9c59ad627 · outbound
Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models Advclip: Downstream-agnostic adversarial examples in multimodal contrastive learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 25ec6ad5-7450-430d-8bc1-36f5a9a20c14 · inbound
Automatic Hard Example Synthesis with Multi-Level Agentic Data Curation Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.