Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2106.13230.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:50:32.514966Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T19:08:49.804199Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation f9c80622-dd4a-4296-8644-c964521a19a2 · inbound
Florence: A New Foundation Model for Computer Vision Video Swin Transformer
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f61f62ae-f6e4-414c-9216-2a8e3757797d · inbound
Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives Video Swin Transformer
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 268ae948-2c9d-46bb-9d91-28360e4305ab · inbound
Rapid Reconstruction of Extremely Accelerated Liver 4D MRI via Chained Iterative Refinement Video Swin Transformer
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 366c63c6-aa71-4bb7-a88c-1f8bda6db2b2 · inbound
Mask-RadarNet: Enhancing Transformer With Spatial-Temporal Semantic Context for Radar Object Detection in Autonomous Driving Video Swin Transformer
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bee6127-0dcc-43df-a926-b519a3dd18a4 · inbound
Visual WetlandBirds Dataset: Bird Species Identification and Behavior Recognition in Videos Video Swin Transformer
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de12811d-cc0d-4bcc-a084-55f0f3072b63 · inbound
Data-Efficient Challenges in Visual Inductive Priors: A Retrospective Video Swin Transformer
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31bac54f-710e-42bd-9f64-fc4ae6435b56 · inbound
Feature Hallucination for Self-supervised Action Recognition Video Swin Transformer
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08074c3b-45e3-4666-b1fb-f80f23217d01 · inbound
Structured Spectral Graph Learning for Anomaly Classification in 3D Chest CT Scans Video Swin Transformer
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33867df1-31d6-4f57-977a-8e13d95fb8cb · inbound
T-MASK: Temporal Masking for Probing Foundation Models across Camera Views in Driver Monitoring Video Swin Transformer
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45a90db6-05b5-4f1c-aa47-26c4cf7f6668 · inbound
Every Subtlety Counts: Fine-grained Person Independence Micro-Action Recognition via Distributionally Robust Optimization Video Swin Transformer
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fca90970-7504-4e69-8218-5e7a1f3c5984 · inbound
Structured Spectral Graph Representation Learning for Multi-label Abnormality Analysis from 3D CT Scans Video Swin Transformer
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dad599ad-5493-4212-86c6-1541f35a956a · inbound
RobustSora: De-Watermarked Benchmark for Robust AI-Generated Video Detection Video Swin Transformer
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b7182760-3804-4eba-a166-9ffc836ea0fc · inbound
Multimodal Anomaly Detection for Human-Robot Interaction Video Swin Transformer
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a2f356bf-7d7a-4703-b1fe-d420fcf31f3f · inbound
ConvFormer3D-TAP: Phase/Uncertainty-Aware Front-End Fusion for Cine CMR View Classification Pipelines Video Swin Transformer
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e485d4af-f542-482a-8768-5f3a1d18db1f · inbound
DVAR: Adversarial Multi-Agent Debate for Video Authenticity Detection Video Swin Transformer
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 58128543-0302-4e1b-b32f-20cc9b8ad66a · inbound
SignMAE: Segmentation-Driven Self-Supervised Learning for Sign Language Recognition Video Swin Transformer
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4d1c6c00-887c-4f63-87b8-1ab91b108f48 · inbound
From Articulated Kinematics to Routed Visual Control for Action-Conditioned Surgical Video Generation Video Swin Transformer
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4fd0229e-93a6-4b99-b806-09946d6b93b6 · inbound
MegaScale-Omni: A Hyper-Scale, Workload-Resilient System for MultiModal LLM Training in Production Video Swin Transformer
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6a28b164-8843-4293-bf48-29a75376b793 · inbound
CAM-VFD: Cross-Attention Multimodal Video Forgery Detection Video Swin Transformer
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5aaa64a0-7d81-4001-be48-c7f848a98c88 · inbound
Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers Video Swin Transformer
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 66797a37-6d28-4a73-99d2-a0a9d73ece4d · inbound
VTI-CoT: Visual-Textual Interleaved Chain of Thought for Video Reasoning Video Swin Transformer
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b2f6fa34-a672-4bb2-8fe8-f559dbbf406b · inbound
MLT-Dedup: Efficient Large-Scale Online Video Deduplication via Multi-Level Representations and Spatial-Temporal Matching Video Swin Transformer
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation aefef589-a3d6-4877-a7fd-db1cbd7fc10e · inbound
Spatio-Temporal Fusion Model for Standard View Classification of Echocardiographic Videos Video Swin Transformer
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c0767247-40c6-4882-a33a-79aed324b7e0 · inbound
Spatio-Temporal Wildfire Spread Prediction in Canada using a Video Swin-Hybrid-U-Net and Satellite Imagery Video Swin Transformer
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c8c7c708-419f-498a-94fa-83116a766446 · inbound
GMoT: Gated Motion-Aware Tokenization for Fine-Grained Micro-Gesture Video Reasoning with Multimodal LLMs Video Swin Transformer
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.