Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:37:19.640172Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2507.02271.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:37:19.640172Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 69a048b8-bf54-4b91-a693-0876059b899b · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation The Foley grail: The art of performing sound for film, games, and animation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1b4e5e8d-32a5-4203-9e59-b2e0858f143d · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Deep residual learning for image recog- nition
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b0c46848-49bc-4f87-9f79-5d9cb000a8a2 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Denoising diffusion probabilistic models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ace39fec-2829-4717-88cf-34db85b6c68e · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Densely connected convolutional networks
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4f00a0c2-f96d-4e06-944f-345201604a8e · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Read, Watch and Scream! Sound Generation from Text and Video
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db35df16-0f1b-47b4-8715-a7105ef1ecd5 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Diverse part dis- covery: Occluded person re-identification with part-aware transformer
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea159ff9-6ce8-4424-849c-83af5bdab436 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Mitigating and evaluating static bias of ac- tion representations in the background and the foreground
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6dce6e9e-f5dc-432a-b213-05cb950ed00d · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation AudioLDM: Text-to-audio generation with latent diffusion models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d62a57df-c626-45fd-b39a-a55d007b256a · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f82e688-1ece-4d83-a2b4-dab92dded7ad · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Diff-foley: Synchronized video-to-audio synthesis with latent diffusion models.Advances in Neural Information Processing Systems, 36,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5212535c-5019-435a-9b49-e5dc42b12f9c · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation The filmmaker’s eye: The language of the lens: The power of lenses and the expressive cinematic image
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 25fe33b3-7429-46af-b274-99c798f89081 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Masked generative video-to-audio transformers with enhanced synchronicity
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e12fde5f-5029-4e60-9cde-02ca65adfc27 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation High-resolution image synthesis with latent diffusion models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e6ed0c00-47cb-4112-a669-9ad94db08c51 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Grad-cam: Visual explanations from deep networks via gradient-based localization
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbcb4214-e9cb-4123-a6c0-3656e012e064 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Denoising diffusion implicit models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a879510a-756e-404b-aee7-f14b9d686219 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Temporally Aligned Audio for Video with Autoregression
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5364271f-4cb6-4cee-88c7-b28e96dcf8dd · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Removing the background by adding the background: Towards background robust self- supervised video representation learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 61ca4cac-0cdc-4c81-b490-83829effb8c5 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Frieren: Efficient Video-to-Audio Generation Network with Rectified Flow Matching
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcf0d7b2-3fe4-4f7c-8ca5-6778b45c9083 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Sonicvisionlm: Playing sound with vi- sion language models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c58c3f2a-5a27-4d6e-a756-6457b59f5e04 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Data- distortion guided self-distillation for deep neural networks
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8a4a7f3c-41ba-4ec9-ba7c-be28db54ae05 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Cutmix: Regularization strategy to train strong classifiers with localizable features
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 13c81013-ec80-41f2-98a5-f4b191318ce7 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Self- supervised scene de-occlusion
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7a4233b5-94f2-40d4-acc6-21b953d901bb · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Be your own teacher: Improve the performance of convolutional neural networks via self distillation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c65f088-3fe7-4bc2-9256-114872a13097 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 990972e0-a7ab-4afc-afc2-e81b7f28873e · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Human orientation estimation un- der partial observation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 085339fe-e831-4b9d-8a2b-553feb5cd84a · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Knowledge distillation by on-the-fly native ensemble
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4e6e21e1-ceb6-42c3-861e-f6b90451c68a · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Vggsound: A large-scale audio-visual dataset
Reference 2014
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ec3772c3-aca5-4cc1-9ee4-a98870fc188b · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Classifier-Free Diffusion Guidance
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60093ecc-6da4-4ccd-9ebf-89a2145e40ac · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Distilling the Knowledge in a Neural Network
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e812d4d2-3a71-4312-be93-55647037a939 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Taming visually guided sound generation
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1d3a90e3-5301-4b14-9610-8063bbac5269 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Pose-guided feature alignment for occluded person re-identification
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 73dbc52d-bd2e-4109-872d-1649253d7afb · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Diffusion models beat gans on image synthe- sis
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 943e0a8e-ed34-4e15-b066-937d0624c7e3 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Motion-aware contrastive video rep- resentation learning via foreground-background merging
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 58f6c621-ff98-454b-abce-0de6cd318e58 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Conditional gener- ation of audio from video via foley analogies
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aaedeca7-da18-4389-996e-139a69818b1f · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Learn2augment: learning to composite videos for data augmentation in ac- tion recognition
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9cfa774b-c41b-4853-96fb-8a535021a89f · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Instance-wise occlusion and depth orders in natural scenes
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bcfd0b05-4dc7-41cc-801b-84536be8fcc6 · outbound
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation STA-V2A: Video-to-Audio Generation with Semantic and Temporal Alignment
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.