Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2109.08472.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T13:31:23.569839Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T20:00:08.259998Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 5f97fecb-928b-4e39-a1ce-cd7485922d02 · inbound
InternVideo: General Video Foundation Models via Generative and Discriminative Learning ActionCLIP: A New Paradigm for Video Action Recognition
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 63c284db-2302-4cfa-87e9-dda14f932cb2 · inbound
SRL-CLIP: Efficient CLIP Video Adaptation via Structured Semantic Role Labels ActionCLIP: A New Paradigm for Video Action Recognition
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1a11ae7e-743f-4e0d-b04e-64b1d6bc76f4 · inbound
LoRA-TTT: Low-Rank Test-Time Training for Vision-Language Models ActionCLIP: A New Paradigm for Video Action Recognition
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 717a7f45-29f6-4f8d-a1c0-9a0d749f5d34 · inbound
SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living ActionCLIP: A New Paradigm for Video Action Recognition
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04503e34-8e11-4846-8e67-8a304247919c · inbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners ActionCLIP: A New Paradigm for Video Action Recognition
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db75da33-556d-40dd-8625-aa56000599e1 · inbound
Conformal Predictions for Human Action Recognition with Vision-Language Models ActionCLIP: A New Paradigm for Video Action Recognition
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64738afd-5232-4775-aa1e-ca4dfe9670b1 · inbound
M2R2: MultiModal Robotic Representation for Temporal Action Segmentation ActionCLIP: A New Paradigm for Video Action Recognition
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ca03ad4d-a30e-4554-a856-6b3bd56dea65 · inbound
From Data to Modeling: Fully Open-vocabulary Scene Graph Generation ActionCLIP: A New Paradigm for Video Action Recognition
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9324a207-f7ac-4fba-aa21-6728672fc579 · inbound
From Motion to Behavior: Hierarchical Modeling of Humanoid Generative Behavior Control ActionCLIP: A New Paradigm for Video Action Recognition
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f133875f-465d-4055-a12b-a1b2bf6f951e · inbound
FRAME: Pre-Training Video Feature Representations via Anticipation and Memory ActionCLIP: A New Paradigm for Video Action Recognition
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db80e471-c663-44e1-999d-d3c94c8a6fbd · inbound
Feature Hallucination for Self-supervised Action Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 168
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fefb174e-7ad2-4264-838b-62483d589634 · inbound
MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e85078f-a7d3-4f7d-9755-7bc0e018c68a · inbound
LLM-enhanced Action-aware Multi-modal Prompt Tuning for Image-Text Matching ActionCLIP: A New Paradigm for Video Action Recognition
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7064ddaa-7c79-407f-98fb-91c5b4832711 · inbound
MReg: A Novel Regression Model with MoE-based Video Feature Mining for Mitral Regurgitation Diagnosis ActionCLIP: A New Paradigm for Video Action Recognition
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e0caaa7-7186-43ea-a47c-e988f75f6f33 · inbound
Teaching Time Series to See and Speak: Forecasting with Aligned Visual and Textual Perspectives ActionCLIP: A New Paradigm for Video Action Recognition
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82048a79-c77a-45fa-830c-fd89501bfcc4 · inbound
"Before, I Asked My Mom, Now I Ask ChatGPT": Visual Privacy Management with Generative AI for Blind and Low-Vision People ActionCLIP: A New Paradigm for Video Action Recognition
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbe83651-a236-407f-91e9-f4fb52d25874 · inbound
Cross-Modal Dual-Causal Learning for Long-Term Action Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e899e0e-25bf-4e3e-ab1b-d6b312477551 · inbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space ActionCLIP: A New Paradigm for Video Action Recognition
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 332235fb-b1dd-4aa6-915b-f080b3dd2d37 · inbound
MoExDA: Domain Adaptation for Edge-based Action Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8793f05a-a5f2-4c71-9e9d-648154d60476 · inbound
Constrained Prompt Enhancement for Improving Zero-Shot Generalization of Vision-Language Models ActionCLIP: A New Paradigm for Video Action Recognition
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d76a287-af6c-469c-bde0-b99ccc3cfd3f · inbound
What Can We Learn from Harry Potter? An Exploratory Study of Visual Representation Learning from Atypical Videos ActionCLIP: A New Paradigm for Video Action Recognition
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d81fe392-c880-4688-9e6e-3676c10a83e5 · inbound
Video Understanding by Design: How Datasets Shape Video Models ActionCLIP: A New Paradigm for Video Action Recognition
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2e75ed0-5b86-48b1-9aaa-87c5ba2f0735 · inbound
Action Hints: Semantic Typicality and Context Uniqueness for Generalizable Skeleton-based Video Anomaly Detection ActionCLIP: A New Paradigm for Video Action Recognition
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0984b6a3-f540-4e40-80e9-ffdbb441b3ec · inbound
Track and Caption Any Motion: Open-Vocabulary Spatiotemporal Captioning via Trajectory-Conditioned Generation ActionCLIP: A New Paradigm for Video Action Recognition
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20e27860-9ce1-44c0-8ec3-f1d915dd61d7 · inbound
Adapting MLLMs for Nuanced Video Retrieval ActionCLIP: A New Paradigm for Video Action Recognition
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de5a1c17-d768-477b-ad83-b51daef083d0 · inbound
Context Matters: Peer-Aware Student Behavioral Engagement Measurement via VLM Action Parsing and LLM Sequence Classification ActionCLIP: A New Paradigm for Video Action Recognition
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 87ef0e5d-2b09-4f9c-93cd-d1c08ccddb87 · inbound
TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c9d0742d-7ef4-4a97-9a98-793f70546892 · inbound
EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges ActionCLIP: A New Paradigm for Video Action Recognition
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2441715e-1b69-4569-845d-2751f516b4fe · inbound
Spatio-Temporal Similarity Volume Aggregation for Open-Vocabulary Action Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 82f2df7e-a194-4303-a9bc-8dd02a225af3 · inbound
VidPrism: Heterogeneous Mixture of Experts for Image-to-Video Transfer ActionCLIP: A New Paradigm for Video Action Recognition
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fbb1ef4b-4342-42aa-b0d0-84eb9d5bcc46 · inbound
Demystifying the Optimal Fair Classifier in Multi-Class Classification ActionCLIP: A New Paradigm for Video Action Recognition
Reference 184
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b89ac15d-cda5-433e-ba95-01e276df524b · inbound
Toward Low-Latency Vision-Language Models with Doubly-Correct Predictions in Egocentric Visual Understanding ActionCLIP: A New Paradigm for Video Action Recognition
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5105fa1e-95b7-4df3-9dbb-c4dc0fe454fa · inbound
TACO: Towards Task-Consistent Open-Vocabulary Adaptation in Video Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f7dcf529-7636-4356-acf1-0530476bda84 · inbound
TACO: Towards Task-Consistent Open-Vocabulary Adaptation in Video Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7b0ae27a-24ac-4b29-8281-5227163d1908 · inbound
In-Context Model Predictive Generation: Open-Vocabulary Motion Synthesis from Language Models to Physics ActionCLIP: A New Paradigm for Video Action Recognition
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 980d0c9c-a66e-42b1-944c-ae7f1d73f149 · inbound
TRUST: Efficient Abdominal Trauma Recognition via Image-to-Ultrasound-Video Transfer Learning ActionCLIP: A New Paradigm for Video Action Recognition
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b1de98bf-1973-4d4e-9c6d-69ee5aaf6fbe · inbound
Breaking the 15% Barrier: A Real-World Data-Driven System for Proactive Social Robot Triggered by User Nonverbal Cues ActionCLIP: A New Paradigm for Video Action Recognition
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71a45313-f943-4d80-943f-5f4f0e7ccf5b · inbound
GHR-VLM: Making Zero-Shot Transit Video Analytics Realizable with Grounded Hybrid Reasoning ActionCLIP: A New Paradigm for Video Action Recognition
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37ad474c-adf1-4974-8d68-3e0b5f0c6071 · inbound
Peak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic Assessment ActionCLIP: A New Paradigm for Video Action Recognition
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cc73b54-d097-46e3-afed-9aa40a8047c3 · inbound
Knowledge-guided Disentanglement with Atomic Actions for Action Recognition ActionCLIP: A New Paradigm for Video Action Recognition
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.