Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2001.08740.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:42:34.185270Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T08:57:47.660611Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 89b87024-bfb2-4ced-b430-40c7b94df63b · inbound
EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models Audiovisual SlowFast Networks for Video Recognition
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e498db1-61bb-4662-9211-c4c727c6e716 · inbound
TOGA: Temporally Grounded Open-Ended Video QA with Weak Supervision Audiovisual SlowFast Networks for Video Recognition
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bd3bc9d-c4ef-470e-ac41-5a75624ca8e4 · inbound
DMAF-Net: An Effective Modality Rebalancing Framework for Incomplete Multi-Modal Medical Image Segmentation Audiovisual SlowFast Networks for Video Recognition
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeeb6297-1c94-4403-b82a-165461958cee · inbound
DEL: Dense Event Localization for Multi-modal Audio-Visual Understanding Audiovisual SlowFast Networks for Video Recognition
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56600a93-ee73-4a74-9978-01870ce512b2 · inbound
Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos Audiovisual SlowFast Networks for Video Recognition
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4f48735-2087-4877-8af9-04073b46d0bd · inbound
OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Audiovisual SlowFast Networks for Video Recognition
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54681c95-aeda-4cf9-aca8-04ec0f6c8b98 · inbound
Attention-Driven Multimodal Alignment for Long-term Action Quality Assessment Audiovisual SlowFast Networks for Video Recognition
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67a76d82-262b-4c59-a43b-2d046a5d4749 · inbound
AViS-Mamba: Adaptive Visual Steering of Audio State-Space Dynamics for Violence Detection Audiovisual SlowFast Networks for Video Recognition
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7cbf2acf-08ef-4009-9fd3-f3db5b8f0b6c · inbound
What-Where Transformer: A Slot-Centric Visual Backbone for Concurrent Representation and Localization Audiovisual SlowFast Networks for Video Recognition
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4809551d-1056-40b8-b724-475f4aee7105 · inbound
USV: Towards Understanding the User-generated Short-form Videos Audiovisual SlowFast Networks for Video Recognition
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d10bf92b-4f89-4da0-bbc0-aa8103feef67 · inbound
SynIB: Informational Bottleneck for Maximizing Synergy in Multimodal Learning Audiovisual SlowFast Networks for Video Recognition
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a8885ba-b003-423e-a10c-cc72a3821ff8 · inbound
On Aligning Hierarchical Standardized Embedding for Audio-visual Generalized Zero-shot Learning Audiovisual SlowFast Networks for Video Recognition
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.