Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:45:08.841789Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2506.20995.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:45:08.841789Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b2efef86-b747-425b-86c4-8cc021f95666 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance The Foley Grail: The Art of Performing Sound for Film, Games, and Animation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bf501168-58c8-491e-88bf-c0087732d0c7 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 138ac0b1-663d-4891-ae36-377b6e6f1b38 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Erasedraw: Learning to insert objects by erasing them from images
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9cc5593b-09cc-4f8e-9f1e-3ef031053a27 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Action2sound: Ambient-aware generation of action sounds from egocentric videos
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f19cb447-815d-43e3-b6e9-057ba58a388e · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Vggsound: A large-scale audio-visual dataset
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 50d1729f-30ce-4e2c-9a55-874ca54bf8db · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Generating visually aligned sound from videos
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 83671cc3-1f8c-46ab-adbb-f180111af7f3 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Video-Guided Foley Sound Generation with Multimodal Controls
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 465ce467-2a71-45bb-94bb-5bd783ab6dd0 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Mmaudio: Taming multimodal joint training for high-quality video-to-audio synthesis
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f6f79a6e-9c2b-497d-9ccb-aa341c3f8f5e · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Clotho: An audio captioning dataset
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c48835b7-ca7e-4973-9b99-a865742c2709 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Compositional visual generation with energy based models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8fcf95f3-9723-4d9e-8778-5fdb8415bf20 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Scaling rectified flow transformers for high-resolution image synthesis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5a745c0a-b954-408e-bd54-993eacde0289 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Sketch2sound: Controllable audio generation via time-varying signals and sonic imitations
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e2701f7c-992a-43d0-bba1-362698e35d76 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Audio set: An ontology and human-labeled dataset for audio events
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8ab04e99-e18d-484c-9a57-60fc2b2b7602 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Imagebind: One embedding space to bind them all
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d2359448-9249-44de-8680-c523e1dad7d8 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Instructme: An instruction guided music edit and remix framework with latent diffusion models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 784a3cab-fd12-4243-b191-259de7ea9baf · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Classifier-free diffusion guidance
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 791d89c6-44a6-45a4-9721-f24578894e16 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Taming visually guided sound generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3f00ebf4-1299-48dc-809f-7d36449e9505 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Synchformer: Efficient synchronization from sparse cues
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 76d83d63-5816-4a20-bcb7-456bbf0e3d3b · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Audioeditor: A training-free diffusion-based audio editing framework
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a6d67126-3422-4ee6-b220-50bc93246462 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Simultaneous music separation and generation using multi-track latent diffusion models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 543f135c-bde9-4a08-907e-32953996f07c · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Analyzing and improving the training dynamics of diffusion models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f0c3b862-20cb-440b-8185-4c5cc2213f44 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Audiocaps: Generating captions for audios in the wild
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ee2e0e48-327f-416e-93d1-9878a97a85f3 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Panns: Large-scale pretrained audio neural networks for audio pattern recognition
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2285d8a5-0e8d-4396-94bf-2d20c51f43b8 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Vintage: Joint video and text conditioning for holistic audio generation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3b64cfc8-9b49-4187-b718-f8eddc76d437 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Gonzalez, Hao Zhang, and Ion Stoica
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8d5aed84-561b-4ff3-b3da-80fc48acb6f0 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Flow matching for generative modeling
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 72f18168-e4f9-43bd-bd23-35e8754133ee · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Compositional visual generation with composable diffusion models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fbac445a-9263-4075-8796-713d4175d007 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Tell what you hear from what you see-video to audio generation through text
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5c27eb4f-97b7-4233-b673-0f8f00d8fd96 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Diff-foley: Synchronized video-to-audio synthesis with latent diffusion models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 79b5cffe-d872-4a6b-a594-afe71466b990 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Multi-source diffusion models for simultaneous music generation and separation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c51d7c4-9601-4119-a8e5-7b6e238e7efa · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Wavcaps: A chatgpt-assisted weakly-labelled audio captioning dataset for audio-language multimodal research
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 20fca138-03ed-415c-958c-0173fb57af3a · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Audio-visual scene analysis with self-supervised multisensory features
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0f77170d-dce8-464e-84cd-b1b0bb67c57f · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Stemgen: A music generation model that listens
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 186907ed-1502-45d5-82a0-79d9c0153b15 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Movie Gen: A Cast of Media Foundation Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 463dd510-0502-4450-9692-436cc26a425b · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Generalized multi-source inference for text conditioned music diffusion models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b96150f7-b8f9-4748-a4de-37565fc32655 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Improved techniques for training gans
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e65fd7f1-2874-4163-a217-7e3bb9e57f15 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Smartmask: context aware high-fidelity mask generation for fine-grained object insertion and layout control
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a4f55f75-e5ed-4ed4-bb7c-d95b5b065c55 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Visually guided sound source separation with audio-visual predictive coding
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4e31c034-6b9a-4f01-93ab-ae7ff2779e8c · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance sd3.5, 2024
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 37145041-e787-421d-b6f3-a1e17e010e53 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Steinmetz and Joshua D
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 03c38e60-10be-4a19-ae61-8e74395b6276 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Add-it: Training-free object insertion in images with pretrained diffusion models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0bc16ca5-f94f-4c75-ba2a-1b4004b179b2 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Liu, Kevin J
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9777209a-2c83-403a-82c9-35f0b3cdcafe · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Temporally aligned audio for video with autoregression
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f6e8b40c-660b-4d33-9552-06e095b8af5c · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance V2a-mapper: A lightweight solution for vision-to-audio generation by connecting foundation models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5ba41fe0-c0b5-41dc-88ec-625b962f1694 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Frieren: Efficient video-to-audio generation network with rectified flow matching
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e0dde940-a687-4083-9f75-0920aa92541e · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Audit: Audio editing by following instructions with latent diffusion models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d4d331b0-7589-46d7-a134-6f921ba9eae9 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Stable diffusion 2.0 and the importance of negative prompts for good results, 2022
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fbf2a7b1-9d47-4e00-a5ab-eed923cef885 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0102a54a-4938-4296-8fd9-88cd98f822af · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Seeing and hearing: Open-domain visual-audio generation with diffusion latent aligners
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6d026ab3-f7b1-4d29-9af6-1d3023bc86fd · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Adding conditional control to text-to-image diffusion models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f86c4cf0-2d82-4547-8a55-783902287549 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b754503b-876d-4787-b826-07e6136fc464 · outbound
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance Visually guided sound source separation using cascaded opponent filter network
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.