Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2505.10917.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:41:00.805264Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T18:45:00.174957Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 3e4f6cf0-67ff-405f-88ce-3095010eaaa4 · inbound
Fine-Grained Zero-Shot Object Detection VISTA: Enhancing Vision-Text Alignment in MLLMs via Cross-Modal Mutual Information Maximization
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67c046c7-93fe-4e25-8faf-bfd9bf4d1412 · inbound
MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning VISTA: Enhancing Vision-Text Alignment in MLLMs via Cross-Modal Mutual Information Maximization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af063bff-0d23-43a7-bdbf-c9a6ab11253d · inbound
Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models VISTA: Enhancing Vision-Text Alignment in MLLMs via Cross-Modal Mutual Information Maximization
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c0ab3302-a89a-4ce4-8e50-5d72fee52e11 · inbound
Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models VISTA: Enhancing Vision-Text Alignment in MLLMs via Cross-Modal Mutual Information Maximization
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation daa69f7a-4758-43b6-8c68-3db61f49861e · inbound
Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition VISTA: Enhancing Vision-Text Alignment in MLLMs via Cross-Modal Mutual Information Maximization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation badd34c8-4d75-4c39-b6c2-9bc6af3ccad1 · inbound
MIRROR: Aligning Semantic Relations from Language to Image via Gromov--Wasserstein VISTA: Enhancing Vision-Text Alignment in MLLMs via Cross-Modal Mutual Information Maximization
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.