Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T19:34:03.414251Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 3 inbound Pith citation observations for arXiv:2501.09972.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T19:34:03.414251Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:36:42.465952Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T00:49:49.179124Z
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 00e8ac5b-22d9-457f-b6ed-053e4633baad · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cadff098-657a-421f-80d3-28d38c8bff45 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f052bbe5-0e7d-4ea7-bf14-9ced26425e68 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions MusicLM: Generating Music From Text
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4696e85-fef2-4782-a263-d5d766faa2ec · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation edf73a6d-e1a9-4e8f-915d-7a840dd1a4ad · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions MIDI-VAE: Modeling Dynamics and Instrumentation of Music with Applications to Style Transfer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e231e4d7-e71a-418f-82a2-1a9143eccb01 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 265024d9-a317-4f79-a9c6-4270aac04a06 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 68a52df8-4290-4d46-a90a-88e412f594a8 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 86c69587-bf6d-4bf3-a99a-2723feed2881 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d7a752d1-77a2-4c28-9574-452d030352ba · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a5065e4-013d-4c2a-b32d-57a51d81548b · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions High Fidelity Neural Audio Compression
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c44e8c3-6c4c-4c34-a6f0-dade0ee9bf5a · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Fast Timing-Conditioned Latent Audio Diffusion
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4752cb0e-7e9e-42ec-927b-42ed9272fb3b · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 32adf641-231b-42a7-b920-f04a5272417b · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions B.; and Torralba, A
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fb0a8788-d61d-47bb-8c29-39e1227e0b4b · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions F.; Ellis, D
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 30ff54a8-c678-4ef1-8ec8-034a48461fe2 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Music Transformer
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dee3761e-d304-46c4-8832-f40997bd36fd · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions MuLan: A Joint Embedding of Music Audio and Natural Language
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74667877-5b8d-48ec-a005-2542e1eb2cfb · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions M$^{2}$UGen: Multi-modal Music Understanding and Generation with the Power of Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3a916e8-be50-411d-93ed-d1cec6538fbb · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions A Comprehensive Survey on Deep Music Generation: Multi-level Representations, Algorithms, Evaluations, and Future Directions
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45bc36ba-4079-46ac-909f-8097602241cf · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a1f01869-c170-4d74-9f71-ce9246a17082 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b114760-cc9e-43f5-93b9-15d7120d3723 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions A.; and Kanazawa, A
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 81186434-3cb3-417c-bc66-13bcddd12ced · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fd3a5470-c1a3-4e73-b2e5-2f6364715a57 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 2b9a702e-5cfa-4dee-bcea-8e9107ccdec3 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Learning Transferable Visual Models From Natural Language Supervision
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c15d592-2566-45c5-a930-3d23e49cec74 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ca21307-eeaf-4a13-8732-059285d8d76b · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3cbf07c3-abb7-4df3-a6cf-a9634488c948 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions V2Meow: Meowing to the Visual Beat via Video-to-Music Generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9ee946ed-7ba9-4bfd-ac61-7804205eddb2 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f8eaa8e8-f8fe-4987-b672-0ca2b7892cc6 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions C.; and Salamon, J
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e409faa0-6990-439b-8bc4-fa809be1e61b · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 463709a4-53bd-4274-87e4-aa20368b5d2c · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions N.; Kaiser, .; and Polosukhin, I
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35656cb0-1090-4d42-9119-aff59f59a1b7 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions NExT-GPT: Any-to-Any Multimodal LLM
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c638963a-d613-4833-8a32-50babbfec5b3 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81dafed8-a95e-4883-a0df-9d34b8bab202 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Long-Term Rhythmic Video Soundtracker
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a956133d-1c98-4525-bbcf-8bdc96119c82 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions SoundStream: An End-to-End Neural Audio Codec
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90dd70bd-6c09-441f-8f38-24e0b0ee568d · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Quantized GAN for Complex Music Generation from Dance Videos
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25c13174-86e4-407e-9422-7f8f0f2e4892 · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Discrete Contrastive Diffusion for Cross-Modal Music and Image Generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5d127102-0d5e-43dd-8c7f-1367dd12317c · outbound
GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8dfead22-de67-4f01-b641-00b7cddd2890 · inbound
AudioGenie: A Training-Free Multi-Agent Framework for Diverse Multimodality-to-Multiaudio Generation GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab03e73b-0eb7-424e-86c1-3955da2955bd · inbound
Video-Guided Text-to-Music Generation Using Public Domain Movie Collections GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 707f5136-b83a-41bc-bac4-8007b047d9ec · inbound
Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections GVMGen: A General Video-to-Music Generation Model with Hierarchical Attentions
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.