Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 33 inbound Pith citation observations for arXiv:2104.08860.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:23:18.768419Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T00:09:15.255843Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 7e639819-48cb-481d-bca9-afea9da9f0da · inbound
Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d552b34c-5ce0-4e56-8486-79dd1ee6230b · inbound
Demystifying CLIP Data CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e45785c1-8ef1-41aa-a520-d9f87f4335e9 · inbound
SRL-CLIP: Efficient CLIP Video Adaptation via Structured Semantic Role Labels CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a0f7aa7c-4b58-4382-9d42-9f1779120058 · inbound
LLaVA-Video: Video Instruction Tuning With Synthetic Data CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 210
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3a1d64b6-eb39-477a-98f1-62d3e1a5c280 · inbound
AstroM$^3$: A self-supervised multimodal model for astronomy CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 544c7b6c-8643-477f-b6c9-ee214541a8b9 · inbound
Whats in a Video: Factorized Autoregressive Decoding for Online Dense Video Captioning CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ea95b71-15e6-45ac-a3b7-82a473567ab1 · inbound
Needle: A Generative AI-Powered Multi-modal Database for Answering Complex Natural Language Queries CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d5d01f8-1b03-425e-bab1-55dc2bbe0c20 · inbound
MADGEN: Mass-Spec attends to De Novo Molecular generation CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92c26dc6-4b83-4553-a6ca-f3662ab335d6 · inbound
Vision-Language Models Do Not Understand Negation CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a36bfaa-22be-4fb5-92c1-7335d9074463 · inbound
CLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b6c544f-d6ff-45c9-ac0f-04e709b8b115 · inbound
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17b3a461-14f7-47f8-a8ff-fbd05e9e0b22 · inbound
Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97af0fe0-60ea-4c3a-b795-75a52cc6c807 · inbound
Exploring Object Status Recognition for Recipe Progress Tracking in Non-Visual Cooking CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 499b8bba-365c-49d0-80ec-fbbeb4520bba · inbound
Regularizing Subspace Redundancy of Low-Rank Adaptation CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0f7ebcc-fc1a-409f-a5c6-ebf1340f57dd · inbound
Video Understanding by Design: How Datasets Shape Video Models CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 243
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f26efacf-4cd4-448b-b021-45248dea36ad · inbound
MSAM: Multi-Semantic Adaptive Mining for Cross-Modal Drone Video-Text Retrieval CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ce8db31-0a39-4f83-822f-0b39afdb2979 · inbound
Adapting MLLMs for Nuanced Video Retrieval CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b2eb13b6-a277-40e2-a154-7f6a4c60b214 · inbound
VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 586e7a31-488a-496e-9a38-222c98c78703 · inbound
Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5a9f84e7-0509-4cf8-a276-f9ec489e4730 · inbound
MP-ISMoE: Mixed-Precision Interactive Side Mixture-of-Experts for Efficient Transfer Learning CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e8c95ba4-e1e7-4552-aedb-8b882c9f613f · inbound
Look Beyond Saliency: Low-Attention Guided Dual Encoding for Video Semantic Search CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 87cfe60d-3ef6-44fe-9c5a-546387cc0ee6 · inbound
Cross-Modal-Domain Generalization Through Semantically Aligned Discrete Representations CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ba1939d0-95c7-4c34-8bab-9027a636fa86 · inbound
Cross-Modal-Domain Generalization Through Semantically Aligned Discrete Representations CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9764b215-1560-48d0-ab4e-fc4cabaa5b5c · inbound
OmniRetriever: Any-to-Any Audio-Video-Text Retrieval via Fusion-as-Teacher Distillation CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d78b6606-e4de-4570-a2a5-daa99bb65ab7 · inbound
Reasoning Text-to-Video Retrieval for Operating Room Clips via Action-Driven Digital Twins CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 52547ef6-7fe1-4c88-a9a0-ab204e70d174 · inbound
LARE: Low-Attention Region Encoding for Text-Image Retrieval CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 76321b28-b6b5-4d49-8a86-e750b1d2fd32 · inbound
VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation eb3295a6-4ac8-4ff7-b996-ca452ecc1e35 · inbound
Video-Text Temporal Localization via Multi-Scale Convolution and Dynamic Routing CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 559a8468-8043-4bd0-ad4b-54f04ec750ab · inbound
Prompting-MammAlps: Fine-Grained Text-to-Video Retrieval for Camera-Trap Data CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 371c535c-6c7c-4f15-b28d-535a3c791952 · inbound
Blurring Modal Boundaries: A Unified Survey from Single- to Multi-Modal Person Re-ldentification CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 153
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 879e6ef0-36cd-460b-ab67-452bc62fe40c · inbound
Trajectory-aware Cross-view Geo-localization with Sequential Observations CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 482c4a35-b131-4ba4-a023-80cb7c0e3616 · inbound
LAVIFT: Latent-Action-Guided Vision Fine-Tuning for Surgical Interaction Recognition CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c406d12-39be-46f6-a23c-b57b58c50c7c · inbound
Knowledge-guided Disentanglement with Atomic Actions for Action Recognition CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.