Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2303.16058.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:32:15.474943Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-17T03:27:59.199947Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ab5c4fc4-b242-4bda-93f8-b7df99450714 · inbound
VideoChat: Chat-Centric Video Understanding Unmasked Teacher: Towards Training-Efficient Video Foundation Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c0869013-6a87-41ef-a2fa-73804324fa1b · inbound
InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation Unmasked Teacher: Towards Training-Efficient Video Foundation Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6f283b83-764f-4bab-9f4d-1a0e81969c34 · inbound
LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment Unmasked Teacher: Towards Training-Efficient Video Foundation Models
Reference 153
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dbb38b48-8acd-4720-92d5-e473c380ada2 · inbound
InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks Unmasked Teacher: Towards Training-Efficient Video Foundation Models
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 20f72546-6c43-4c18-be9a-2fe99573e7d9 · inbound
Revisiting Feature Prediction for Learning Visual Representations from Video Unmasked Teacher: Towards Training-Efficient Video Foundation Models
Reference 188
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9a1be07f-ca20-4d33-941e-20cf47ff3536 · inbound
SMART-Vision: Survey of Modern Action Recognition Techniques in Vision Unmasked Teacher: Towards Training-Efficient Video Foundation Models
Reference 160
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bd19b2e-3c20-4bcf-95e5-c67e41ab1844 · inbound
A Survey on Efficiency Optimization Techniques for DNN-based Video Analytics: Process Systems, Algorithms, and Applications Unmasked Teacher: Towards Training-Efficient Video Foundation Models
Reference 141
Source-reported events for the cited work
Unavailable: canonical work link unavailable.