Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2402.00357.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:31:25.694237Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T01:27:30.991837Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c6eee69c-3885-4fef-b765-55b7dd97713e · inbound
Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback Safety of Multimodal Large Language Models on Images and Texts
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 001a728f-c1d1-479f-9d59-743bc432ad1a · inbound
RapGuard: Safeguarding Multimodal Large Language Models via Rationale-aware Defensive Prompting Safety of Multimodal Large Language Models on Images and Texts
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b17aab4-b0ab-4e96-8c3f-41bdf1be6cf6 · inbound
Spot Risks Before Speaking! Unraveling Safety Attention Heads in Large Vision-Language Models Safety of Multimodal Large Language Models on Images and Texts
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c919b9d8-7b51-4ed8-9944-1158ad58de36 · inbound
Robust-LLaVA: On the Effectiveness of Large-Scale Robust Image Encoders for Multi-modal Large Language Models Safety of Multimodal Large Language Models on Images and Texts
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd9d8eb5-393c-4e58-a585-61fa9825b70d · inbound
When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs Safety of Multimodal Large Language Models on Images and Texts
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f71db071-708a-45be-900d-a66b1762284c · inbound
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations Safety of Multimodal Large Language Models on Images and Texts
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba15c73c-8609-4fb2-a3be-41e2bb43b959 · inbound
Mapping User Trust in Vision Language Models: Research Landscape, Challenges, and Prospects Safety of Multimodal Large Language Models on Images and Texts
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cccdc95a-89aa-4648-9e47-aaf4bb2cc0e4 · inbound
Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment Safety of Multimodal Large Language Models on Images and Texts
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fae58dc4-feec-41ee-ac7a-2ab4edfe5819 · inbound
Robustness Evaluation of OCR-based Visual Document Understanding under Multi-Modal Adversarial Attacks Safety of Multimodal Large Language Models on Images and Texts
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2bbf271-6201-44be-a588-6223836e8b4e · inbound
The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair Safety of Multimodal Large Language Models on Images and Texts
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbad9768-587a-4bb5-8ba6-3e50a50fa5e4 · inbound
Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Safety of Multimodal Large Language Models on Images and Texts
Reference 146
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3894a040-3054-4e18-a724-87ac4b6121f9 · inbound
Is GPT-4o mini Blinded by its Own Safety Filters? Exposing the Multimodal-to-Unimodal Bottleneck in Hate Speech Detection Safety of Multimodal Large Language Models on Images and Texts
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97a40a27-a51b-49e8-b125-c291234dd07f · inbound
Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing Safety of Multimodal Large Language Models on Images and Texts
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a3431651-ae9c-4ad6-9d92-b0bb7f6ad56a · inbound
Investigating Adversarial Robustness of Multi-modal Large Language Models Safety of Multimodal Large Language Models on Images and Texts
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 339e86b5-919c-4603-ab5f-e6107eca8397 · inbound
Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges Safety of Multimodal Large Language Models on Images and Texts
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d54a7677-0c02-45c0-a740-7977675c926f · inbound
V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure Safety of Multimodal Large Language Models on Images and Texts
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 718fda89-e059-4c42-8d1f-89c3c1e8b2bc · inbound
How China-Origin Vision-Language Models Move from Refusal to Reframing in State Alignment Safety of Multimodal Large Language Models on Images and Texts
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.