Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2402.04252.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:21:41.212978Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 761d635e-cc0e-4563-abd0-2d153a7b3767 · inbound
Perception Encoder: The best visual embeddings are not at the output of the network EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 130
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 33163676-9942-43da-96e7-7883f6a5b377 · inbound
Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6141965c-c4e6-409f-9782-80e21e49fed8 · inbound
mRAG: Elucidating the Design Space of Multi-modal Retrieval-Augmented Generation EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ac52ac6-a3b7-487b-a43c-42578b84fa9e · inbound
Improve Multi-Modal Embedding Learning via Explicit Hard Negative Gradient Amplifying EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b287334b-6846-47e4-9a02-02e8a143c3d6 · inbound
Can Argus Judge Them All? Comparing VLMs Across Domains EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d1c2b4e-a74b-4f70-8c1c-f0b2b2932a61 · inbound
Temporal Chain of Thought: Long-Video Understanding by Thinking in Frames EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e092c2f-2333-4d1d-849b-1b18121ed4e0 · inbound
ExpStar: Towards Automatic Commentary Generation for Multi-discipline Scientific Experiments EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5006196a-e851-43d4-b838-bb79c4440a6c · inbound
Category-level Text-to-Image Retrieval Improved: Bridging the Domain Gap with Diffusion Models and Vision Encoders EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5a8c8bd-7223-4a3f-a987-fd10e00d65c9 · inbound
Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c74decde-c081-47ff-baf8-5c04afc10656 · inbound
QKVQA: Question-Focused Filtering for Knowledge-based VQA EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24b1a7b5-fabe-48c9-99c0-5ccc3d4e1fe9 · inbound
Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 210dbf3f-69e6-41c3-ba71-1ed8a84f9b3b · inbound
Explicit Logic Channel for Validation and Enhancement of MLLMs on Zero-Shot Tasks EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a43f5ecc-af12-4c0e-b663-17de2f84e7cd · inbound
Explicit Logic Channel for Validation and Enhancement of MLLMs on Zero-Shot Tasks EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab31082e-b61d-4101-a65d-f26a4078c97f · inbound
CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9dd7262-0e5a-42be-91d3-07b0c2fb289d · inbound
MG$^2$-RAG: Multi-Granularity Graph for Multimodal Retrieval-Augmented Generation EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 58593182-a8bb-40e1-84b2-f315e94771f7 · inbound
MG$^2$-RAG: Multi-Granularity Graph for Multimodal Retrieval-Augmented Generation EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f1319da-4fef-4b53-9cca-409ce23938c0 · inbound
Benchmarking Deflection and Hallucination in Large Vision-Language Models EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e3b25e8-7e0a-4d90-8e9a-b9b75bd515fe · inbound
Chain-of-Models Pre-Training: Rethinking Training Acceleration of Vision Foundation Models EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f31a42a9-3475-4e02-aa22-a9e3b35373a2 · inbound
Exploring High-Order Self-Similarity for Video Understanding EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ec127f8-c038-4205-9379-8c004079cb78 · inbound
HiCrew: Hierarchical Reasoning for Long-Form Video Understanding via Question-Aware Multi-Agent Collaboration EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1219347e-580a-4b29-802e-2749de453990 · inbound
Sensorimotor World Models: Perception for Action via Inverse Dynamics EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f7e01b75-078c-4ca5-8725-73c52626a467 · inbound
ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7bafcb59-91f1-40f5-bec6-abf94911e215 · inbound
Ground Then Rank: Revisiting Knowledge-Based VQA with Training-Free Entity Identification EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3c2d5fa3-575a-4d5f-a188-1c1710b824b6 · inbound
ReasonCLIP-58M: Visually Grounded Commonsense Reasoning Supervision for CLIP EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38118793-1777-41f9-b8f6-6a1acc7ce684 · inbound
Identifying and Resolving Pitfalls of Knowledge-Based VQA Benchmarks: Auditing, Repairing, and Augmenting EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 851cac4c-9e90-4435-b87d-30ba397117c2 · inbound
UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 353c7071-a8c2-4dbc-9acf-c8a45a9ce76c · inbound
UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.