Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:05:12.150659Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2505.20038.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:05:12.150659Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
25 of 25 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5001ff69-54ec-4cfe-8d99-1f5e3c17a8b5 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Video-Guided Foley Sound Generation with Multimodal Controls
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bade64f8-cec2-4ab2-ab56-2e7f85d2b730 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2309e642-f146-464a-a083-36fb2b903a66 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d41f6e1-c5b0-4aba-8b10-537ae6f9645d · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26bf0cc3-e0af-410d-8e83-5b219d1d5a88 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Gans trained by a two time-scale update rule converge to a local nash equilib- rium
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be61c9ec-bfb7-4bdc-9141-01856861c8e0 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Taming visually guided sound generation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0562a702-3591-42da-91a1-9f6291dbbe66 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Sophia Koepke, Olivia Wiles, Yael Moses, and Andrew Zisserman
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 84250236-4a2f-410b-8b7e-80c0110144ff · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Crandall, and Christopher Raphael
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8e7bcbcc-767a-40d2-ae68-369c5ecf4d75 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Tri-Ergon: Fine-grained Video-to-Audio Generation with Multi-modal Conditions and LUFS Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 58c5112c-09e0-4fee-8b63-9ea593e6c88e · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Diff-foley: Synchronized video-to-audio synthesis with la- tent diffusion models, 2023
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 142adb7a-b282-4e1f-b92e-0b5caa682562 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Foleygen: Visually-guided audio generation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5726a169-6270-466d-b1fa-5fbad09bc8d6 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Qwen2.5 technical report, 2025
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 571820b7-73f0-4787-ac9a-5534892b862e · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Direct preference optimization: Your language model is secretly a reward model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c40d3e8-45a1-4301-9083-8ca8d04573a2 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Improved techniques for training gans
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fde6f11-e80f-45b7-9201-2879022d1fe5 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Audeo: Audio Generation for a Silent Performance Video
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bfde4dd2-da87-4953-a989-e1eb11484bd0 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks AudioX: A Unified Framework for Anything-to-Audio Generation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8aab714f-f38e-46bd-abbc-3c3d4a717b46 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Temporally Aligned Audio for Video with Autoregression
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4163fa1e-cbdb-4479-8a77-037d6e2d1217 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks V2a-mapper: A lightweight solution for vision-to-audio generation by connecting foun- dation models, 2023
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c847c1af-9770-42ea-b268-32cd39ee6b86 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Frieren: Efficient Video-to-Audio Generation Network with Rectified Flow Matching
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d01e4aff-f9f4-49a1-abdd-c810a9e3ff68 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Chain-of-thought prompting elicits reasoning in large lan- guage models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04a16f20-97f0-4ae9-ba51-3c4b66cfaf48 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Large-scale con- trastive language-audio pretraining with feature fusion and keyword-to-caption augmentation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c82a12f4-a32d-4160-8063-9d5fe817a8ea · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks LLaVA-CoT: Let Vision Language Models Reason Step-by-Step
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 408206cd-881b-4eb6-be8a-b9c0006f1583 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Diverse and aligned audio-to-video genera- tion via text-to-video model adaptation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 10a5b411-b034-47a1-b8e3-b003ce232157 · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Improve Vision Language Model Chain-of-thought Reasoning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5935e36-2c37-445e-81ae-9d1e08b05f4c · outbound
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.