Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2102.05095.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:55:28.412053Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1359
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 140aac23-27fa-41c3-8ea3-ac9d6efb6e82 · inbound
Video Diffusion Models Is Space-Time Attention All You Need for Video Understanding?
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ab418a0-6bfc-4f97-990e-a8b928f6f186 · inbound
MOOSE: Pay Attention to Temporal Dynamics for Video Understanding via Optical Flows Is Space-Time Attention All You Need for Video Understanding?
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeb68259-9d9e-44d4-b995-f92cc6f0856e · inbound
Fine-Tuning Video Transformers for Word-Level Bangla Sign Language: A Comparative Analysis for Classification Tasks Is Space-Time Attention All You Need for Video Understanding?
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e0c1bef-2930-49ac-872f-316afaa219da · inbound
Data-Efficient Challenges in Visual Inductive Priors: A Retrospective Is Space-Time Attention All You Need for Video Understanding?
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a890a18d-7361-412a-8960-615401f929aa · inbound
Vision Transformer-Based Time-Series Image Reconstruction for Cloud-Filling Applications Is Space-Time Attention All You Need for Video Understanding?
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd7e80c6-f1a6-4542-be12-6e0b683c42a7 · inbound
Comparing Learning Paradigms for Egocentric Video Summarization Is Space-Time Attention All You Need for Video Understanding?
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c41e3d9-773f-4c9d-bf1e-460f39b99941 · inbound
MVP: Winning Solution to SMP Challenge 2025 Video Track Is Space-Time Attention All You Need for Video Understanding?
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9c232d4-c6b1-404e-8f17-5e6da800d8ae · inbound
Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes Is Space-Time Attention All You Need for Video Understanding?
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82bce4bf-2976-4f3a-a159-93b903fc27f8 · inbound
CascadeFormer: A Family of Two-stage Cascading Transformers for Skeleton-based Human Action Recognition Is Space-Time Attention All You Need for Video Understanding?
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a5d2dfb-a2a1-4b8a-bbeb-84eca88c0abd · inbound
Cataract-LMM Large-Scale Multi-Source Multi-Task Benchmark for Deep Learning in Surgical Video Analysis Is Space-Time Attention All You Need for Video Understanding?
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76e0ead7-bfbe-4fa7-9d3f-13c125b192eb · inbound
A Space-Time Transformer for Precipitation Nowcasting Is Space-Time Attention All You Need for Video Understanding?
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45670124-a771-4715-b6df-cf104cb965f5 · inbound
RobustSora: De-Watermarked Benchmark for Robust AI-Generated Video Detection Is Space-Time Attention All You Need for Video Understanding?
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0c425cb9-266e-4991-aa6e-391b8754c243 · inbound
Explainable Fall Detection for Elderly Monitoring via Temporally Stable SHAP in Skeleton-Based Human Activity Recognition Is Space-Time Attention All You Need for Video Understanding?
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8fd48442-8300-4968-afb9-7ebebff1f7dc · inbound
Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction Is Space-Time Attention All You Need for Video Understanding?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a4b40fe-1be0-441a-bd1e-de4a4bf32084 · inbound
Exploring High-Order Self-Similarity for Video Understanding Is Space-Time Attention All You Need for Video Understanding?
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 539c8429-561d-4ffc-b7fa-1f728c33144f · inbound
Only Brains Align with Brains: Cross-Region Alignment Patterns Expose Limits of Normative Models Is Space-Time Attention All You Need for Video Understanding?
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d788f70-2247-45f5-90fb-baff01f58a14 · inbound
Parameter-Efficient Multi-View Proficiency Estimation: From Discriminative Classification to Generative Feedback Is Space-Time Attention All You Need for Video Understanding?
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e67b4d71-56db-49d5-85a0-7d6e499a8a29 · inbound
LookWhen? Fast Video Recognition by Learning When, Where, and What to Compute Is Space-Time Attention All You Need for Video Understanding?
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation edb86348-ec6e-48fb-9f45-737820589425 · inbound
CAM-VFD: Cross-Attention Multimodal Video Forgery Detection Is Space-Time Attention All You Need for Video Understanding?
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a2303355-69b9-4e82-a690-322acc4b77b5 · inbound
Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers Is Space-Time Attention All You Need for Video Understanding?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation daa30a4d-0e4f-4095-b886-38e9453a7445 · inbound
Signed Dual Attention: Capturing Signed Dependencies in Time Series Forecasting Is Space-Time Attention All You Need for Video Understanding?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 681dc00c-6bde-453d-a614-ff4caba6f2f5 · inbound
A multi-task spatiotemporal deep neural network for predicting penetration depth and morphology in laser welding Is Space-Time Attention All You Need for Video Understanding?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 520529b3-782c-46a2-a61c-cdd8a937fe75 · inbound
Incentivizing Vision Language Models to Search for Long Video Question Answering Is Space-Time Attention All You Need for Video Understanding?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93a2de9b-9a1a-4845-874b-f4b23582ace7 · inbound
The evolution of AI from image interpretation toward scientific inference in nanoparticle electron microscopy Is Space-Time Attention All You Need for Video Understanding?
Reference 104
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba800377-3a00-4419-9e01-e6dda56e0606 · inbound
GMoT: Gated Motion-Aware Tokenization for Fine-Grained Micro-Gesture Video Reasoning with Multimodal LLMs Is Space-Time Attention All You Need for Video Understanding?
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.