Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2210.07839.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:27:56.939307Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T05:56:39.834913Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation dd674aa9-a0f8-43e6-a43c-906b034b9baa · inbound
The Sound of Water: Inferring Physical Properties from Pouring Liquids Contrastive Audio-Visual Masked Autoencoder
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cb8e19a-25b3-495a-91d9-2549a660af15 · inbound
KDC-MAE: Knowledge Distilled Contrastive Mask Auto-Encoder Contrastive Audio-Visual Masked Autoencoder
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 395dd764-c639-47b6-b6af-fccc845ea6ce · inbound
A Survey of Recent Advances and Challenges in Deep Audio-Visual Correlation Learning Contrastive Audio-Visual Masked Autoencoder
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a983b572-ac08-48bc-a575-5ef48fcf5c18 · inbound
TACO: Training-free Sound Prompted Segmentation via Semantically Constrained Audio-visual CO-factorization Contrastive Audio-Visual Masked Autoencoder
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a89a43ff-8a4f-459a-8418-c6a3b65d07b2 · inbound
JoVALE: Detecting Human Actions in Video Using Audiovisual and Language Contexts Contrastive Audio-Visual Masked Autoencoder
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b5530ac-5d7b-4943-9ce5-9a3b7aae252e · inbound
Mitigating Audiovisual Mismatch in Visual-Guide Audio Captioning Contrastive Audio-Visual Masked Autoencoder
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac127a94-68f0-4b79-b5e2-1c199391e695 · inbound
Let Your Video Listen to Your Music! Contrastive Audio-Visual Masked Autoencoder
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8a4800c-84d1-4295-8a4d-20dd0bc26f7b · inbound
Hear-Your-Click: Interactive Object-Specific Video-to-Audio Generation Contrastive Audio-Visual Masked Autoencoder
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8696b57a-3e26-4303-8f96-eccab980b3f1 · inbound
Dynamic Inter-Class Confusion-Aware Encoder for Audio-Visual Fusion in Human Activity Recognition Contrastive Audio-Visual Masked Autoencoder
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abba1354-4e34-4433-88c0-93456f3e50af · inbound
DualDub: Video-to-Soundtrack Generation via Joint Speech and Background Audio Synthesis Contrastive Audio-Visual Masked Autoencoder
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90dd38b2-8ab7-4064-bc64-7caa43145675 · inbound
Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos Contrastive Audio-Visual Masked Autoencoder
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b1f7c7a-b55a-4d1e-b348-b32c898633d5 · inbound
Revisiting Audio-language Pretraining for Learning General-purpose Audio Representation Contrastive Audio-Visual Masked Autoencoder
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85d2e9d4-b713-46dd-8943-ff6ff105517f · inbound
Audio-Visual Continual Test-Time Adaptation without Forgetting Contrastive Audio-Visual Masked Autoencoder
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8c91a3a-8934-4f4c-83ae-7ef929538a5f · inbound
ControlFoley: Unified and Controllable Video-to-Audio Generation with Cross-Modal Conflict Handling Contrastive Audio-Visual Masked Autoencoder
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ab8cf168-448e-4b94-97af-331555ecf050 · inbound
From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning Contrastive Audio-Visual Masked Autoencoder
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0f64479d-064e-458f-b908-30950fa85215 · inbound
FATE: Frame-Level Audio-Visual Temporal Embedding Contrastive Audio-Visual Masked Autoencoder
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a3c7be1-1c5c-46b4-a022-e8d7256a9cfa · inbound
HarmoniDPO: Video-guided Audio Generation via Preference-Optimized Diffusion Contrastive Audio-Visual Masked Autoencoder
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.