Pith. sign in

Paper Citation Record · LEDGER

MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2410.12957.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.12957 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:36:42.531770Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T13:28:18.797278Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a68545f4-b077-4f13-a4aa-7ddfbe06bd97 · inbound

VidMusician: Video-to-Music Generation with Semantic-Rhythmic Alignment via Hierarchical Visual Features cites this paper.

VidMusician: Video-to-Music Generation with Semantic-Rhythmic Alignment via Hierarchical Visual Features MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:27.207535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:27.207535Z digest=sha256:8dbab2d4787176f95b7e894c8333ee2b21fe8f7f0d9d4aee99b0d84bc620bf08

Observation ccc2f29f-7925-4ba2-853c-5c31c45ebc5a · inbound

AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation cites this paper.

AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T11:38:08.379292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:38:08.379292Z digest=sha256:62d766db1d0849bd9c03dbb8b5bd79801c7a3ecbf4863de8053b86c217e9898d

Observation f64a50f9-8283-4b54-a678-867818026ec2 · inbound

AudioGenie: A Training-Free Multi-Agent Framework for Diverse Multimodality-to-Multiaudio Generation cites this paper.

AudioGenie: A Training-Free Multi-Agent Framework for Diverse Multimodality-to-Multiaudio Generation MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:55.658153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:21:55.658153Z digest=sha256:10c744a708685f82148486cb4fe3ceefa426f603360e754796093118f70a0f68

Observation 92882141-7d2d-423f-9e38-e178bfac7bf1 · inbound

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections cites this paper.

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:49:43.716277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:49:43.716277Z digest=sha256:7bdea18db7d6221111420c98eaefa6bd8dda15051501db584bb7055f76fe8611

Observation 570ab3ab-09c6-4f44-a80b-eb1002082a44 · inbound

EXPOTION: Facial Expression and Motion Control for Multimodal Music Generation cites this paper.

EXPOTION: Facial Expression and Motion Control for Multimodal Music Generation MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T19:40:03.310979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:40:03.310979Z digest=sha256:789131dd798a6b3f33314af9676ce9315247c09f232de97f5934b2ff1cc4f889

Observation b1c89c59-1a2a-4c02-bf83-afc3c1a059e2 · inbound

Controllable Video-to-Music Generation with Multiple Time-Varying Conditions cites this paper.

Controllable Video-to-Music Generation with Multiple Time-Varying Conditions MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T13:31:11.375359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:31:11.375359Z digest=sha256:94a2a65d3c87fadac1fda1d1fad28b304da3a774ffb058cfd2623b3d9b9c7351

Observation 00bede47-b91c-4548-92a9-1b49e45abfc6 · inbound

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions cites this paper.

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T00:56:24.684395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T13:10:29.213917Z digest=sha256:84525e2d5bf8d157bccab8927ea30060ffdd8b4234485397c20c0da51d85bc7a

Observation afaac524-bfe9-4f50-8cb7-1b4c376d4569 · inbound

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation cites this paper.

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:28:18.798650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T08:04:48.283908Z digest=sha256:78085413d4d9da7108284a236e80ff57d62b354628f14f7d8c904ddaf47e3532

Observation f06b666f-bace-4d58-a581-5d524fbabd79 · inbound

Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections cites this paper.

Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T00:36:42.531770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:36:42.531770Z digest=sha256:7bd8c5024bf427ef6ebdd65cc230a045898f771eb5c5e3a53dac053f43ae8574