Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:36:42.599014Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2608.11576.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:36:42.599014Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:36:42.391185Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-16T00:36:42.953731Z
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9994dfab-e618-4b92-91d6-f403c9decf4b · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 49f7049e-e796-46e3-9af6-b670310f557f · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Early work in modern video-to-music generation systems uses large-scale web music-video corpora and autoregressive modeling over semantic acoustic tokens to generate music [8]
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d1f0f9bd-1d9e-4177-aebe-5ce875711696 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections trance music
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f0743c4e-cba8-4a45-915f-35bb08426abd · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1369da0c-923d-4a28-b521-da386bbf7aa4 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections First, we bench- mark video-to-music generation tasks by comparing existing video-to-music generation models trained and evaluated on identical data, using OSSL-v2
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 442247b2-c215-4420-b9d4-a54b69ada218 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections +Dialogue
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4d97db7e-2296-4f1a-a2c9-6fff5d6fbc44 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Be- cause the dataset is free from link rot and does not require separate web scraping, our dataset is suitable as a durable benchmark for the field
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation df5e0c5b-906b-47e3-959a-cc4b271b19e6 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Teaser Generation for Long Documentaries and Educational Videos
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 48d25803-3e24-4035-bc77-87021a8676c0 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Attendaffectnet–emotion pre- diction of movie viewers using multimodal fusion with self-attention,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cefa76f8-82df-4414-a38f-80e911720855 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Predicting emotion from music videos: exploring the relative contribution of visual and auditory information to affective responses
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aefa4075-821b-4a94-9155-3c07c6c43813 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections The cognitive processing of film and musical soundtracks,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 008a119a-ccf9-473b-b194-b9633b4b2f70 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Multimodal deep models for predicting affec- tive responses evoked by movies.,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fa33de3d-8bb4-4105-b5cc-52ed53cdaad7 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections On music’s potential to convey meaning in film: A systematic review of empirical evi- dence,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e4bb1dcc-a488-4e3f-b982-43011b6020d4 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Emotion Embedding Spaces for Matching Music to Stories
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b24fc0fe-dfd4-463c-96d2-b5d1293e74d7 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Foley music: Learning to generate music from videos,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4b2cefab-443c-4b9e-8172-064e5352d578 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections V2meow: Meow- ing to the visual beat via video-to-music generation,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fd6a3ed6-191b-4b45-8089-1bf9bbd4218a · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Vidmuse: A simple video-to-music generation framework with long-short-term modeling,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1ceb4568-70ff-40a4-9f39-ff64d417ccc9 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Vmas: Video-to-music generation via semantic alignment in web music videos,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0a616f9c-5105-4098-9803-76aed2621f3b · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Sonique: Video background music generation using unpaired audio- visual data,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 707f5136-b83a-41bc-bac4-8007b047d9ec · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f194b0-b420-4d16-8e89-42189f1c22a4 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Vision-to-Music Generation: A Survey
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a329da2a-9f37-45c7-a4d3-bad6ee9ea753 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Ai- based chinese-style music generation from video con- tent: a study on cross-modal analysis and generation methods,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 81b7990b-37ce-403d-83d9-2d9c84b6d211 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections V2M-Zero: Zero-Pair Time-Aligned Video-to-Music Generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation aa944ece-843f-4cd3-9506-b9cc63eb9cc7 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Diff- v2m: A hierarchical conditional diffusion model with explicit rhythmic modeling for video-to-music genera- tion,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 10022ec1-ca19-4aa1-bf0f-aad20ccbb1fe · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Quan- tized gan for complex music generation from dance videos,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b2d7c4b9-30bd-4184-86b6-a939bf58ac3e · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Video background music genera- tion: Dataset, method and evaluation,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 990607ec-3e49-4516-874e-35c7c1f52409 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Content-Based Video-Music Retrieval Using Soft Intra-Modal Structure Constraint
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8147f160-bd41-4192-82f0-c30b73745ab6 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Video-Guided Text-to-Music Generation Using Public Domain Movie Collections
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa6e8aab-9e17-454e-b054-7b958e060045 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Acoustic profiles in vocal emotion expression.,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6abfe8f2-f194-4e2c-84c1-41efdb5a99a1 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Background ducking to produce esthetically pleasing audio for tv with clear speech,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation bee2f69a-b172-452f-ae3e-7047d87f4192 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Improving dialogue intelligi- bility in streaming media,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 120bc34a-93c5-4792-90e6-0256dc84d0c3 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Creating a multitrack clas- sical music performance dataset for multimodal mu- sic analysis: Challenges, insights, and applications,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ebc889d5-7b93-4f4d-9a5e-3aeb1039d244 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Video2music: Suitable music generation from videos using an affective multimodal transformer model,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e1e7e50f-29fd-4335-a102-3ddf5172aec5 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Extending Visual Dynamics for Video-to-Music Generation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd94a98a-47b9-4228-ba40-0b7c2eac11fb · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections VidMusician: Video-to-Music Generation with Semantic-Rhythmic Alignment via Hierarchical Visual Features
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f06b666f-bace-4d58-a581-5d524fbabd79 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e1b4654-799c-4717-9f79-8ce89d3db55f · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Video echoed in music: Semantic, temporal, and rhythmic alignment for video-to-music generation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7c0c2411-6d68-43cc-8757-d2ece80bb540 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 21d31ded-83a0-4005-be0e-91e719cd9342 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Harmonizing Pixels and Melodies: Maestro-Guided Film Score Generation and Composition Style Transfer
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1209e8e0-5c05-4638-a493-3f2d2ff08163 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections FilmComposer: LLM-Driven Music Production for Silent Film Clips
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72562688-b4de-4467-b3ec-8b104b63b241 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections M$^{2}$UGen: Multi-modal Music Understanding and Generation with the Power of Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16bf29af-c87e-4b33-9971-4c33d6f8d77d · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d8c3492-5c3a-4909-9fa0-35fc7b2ed464 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Multimodal Music Generation with Explicit Bridges and Retrieval Augmentation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b074ce61-01be-4387-8b04-d49dfe2f24fd · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Benchmarks and leaderboards for sound demixing tasks
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation eb3d6160-8e0d-4e0f-a8db-8702c047d2af · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Panns: Large- scale pretrained audio neural networks for audio pat- tern recognition,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 653a13bf-c98b-4d2a-83cf-051647f1e180 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Film: Visual reasoning with a general conditioning layer,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da8bf902-2b3d-4bbe-9776-dc74b80f2ac7 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Simple and controllable music generation,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3ce23509-f9dd-4bc1-aec2-972436a940c0 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Learning transferable visual models from natural lan- guage supervision,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3eee001c-eab3-4d51-b092-e28182dd0b42 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Blip-2: Bootstrapping language-image pre- training with frozen image encoders and large language models,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8dba42b6-6f13-4d4f-ac83-62e87cb04ec0 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Stable audio open,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 845fcd3d-3188-46eb-b21f-4d82c9bd4ff7 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 033875e8-7627-4d8c-9f7d-d2ee6ffd8ca0 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Reliable fidelity and diversity metrics for generative models,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 94ed3f79-cf9d-4933-9880-15011d8c22bb · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Large-scale contrastive language-audio pretraining with feature fu- sion and keyword-to-caption augmentation,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8a87dcdc-06b6-4acc-8217-f75e2e86a756 · outbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Efficient Training of Audio Transformers with Patchout
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9994dfab-e618-4b92-91d6-f403c9decf4b · inbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.