Pith. sign in

Paper Citation Record · LEDGER

Video-to-Audio Generation with Hidden Alignment

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2407.07464.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.07464 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:43:41.258788Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:56:39.412382Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 07b7ac75-c128-441c-a497-5f555665cb84 · inbound

Movie Gen: A Cast of Media Foundation Models cites this paper.

Movie Gen: A Cast of Media Foundation Models Video-to-Audio Generation with Hidden Alignment

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:26.132308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T14:16:18.521699Z digest=sha256:66c6de566c7f2f48fe7ff0958ae9e6ea2dc88fb3340370bb47110e33b60c19f5

Observation 0991ced3-6aa9-450d-a670-69c14db4695a · inbound

Video-Guided Foley Sound Generation with Multimodal Controls cites this paper.

Video-Guided Foley Sound Generation with Multimodal Controls Video-to-Audio Generation with Hidden Alignment

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-12T12:00:01.772811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:00:01.772811Z digest=sha256:d347697177f536dce1e9e5fdd16238d957388dd36b72442a5b7a5c60fcf04a3f

Observation 42b465a8-d2b4-4bc1-aed8-b55bee8330f7 · inbound

OmniAudio: Generating Spatial Audio from 360-Degree Video cites this paper.

OmniAudio: Generating Spatial Audio from 360-Degree Video Video-to-Audio Generation with Hidden Alignment

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T11:43:41.247637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:43:41.247637Z digest=sha256:11415115f05bf483ac8e374912401a539f415a5f4e748a1dda8d62aaee31bc14

Observation f7ff4e2d-b1c0-453b-87e0-e1e9f00c598e · inbound

OmniAudio: Generating Spatial Audio from 360-Degree Video cites this paper.

OmniAudio: Generating Spatial Audio from 360-Degree Video Video-to-Audio Generation with Hidden Alignment

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T11:43:41.258788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:43:41.258788Z digest=sha256:d3ddb79f2bd6876693d6bf470fe8d4cd09b3ce6b8ad6df6078405b99403a1243

Observation 9ae3ddbc-b652-4a76-839e-9901d10ec1b6 · inbound

Hearing from Silence: Reasoning Audio Descriptions from Silent Videos via Vision-Language Model cites this paper.

Hearing from Silence: Reasoning Audio Descriptions from Silent Videos via Vision-Language Model Video-to-Audio Generation with Hidden Alignment

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:30.995667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:25:30.995667Z digest=sha256:a396ecffb9ddd6f2d75aa1804607475539443943537519e82f0b76ee313fea96

Observation 1c89f49b-6dd4-4868-8384-4f148663444a · inbound

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance cites this paper.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Video-to-Audio Generation with Hidden Alignment

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.814761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:0bcb01eafaa140296f8e0d86424e5cf3a30a6fc4dda249dca72e023389516935

Observation 74aec35b-5806-40b6-8d49-c687d54f8bac · inbound

AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance cites this paper.

AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance Video-to-Audio Generation with Hidden Alignment

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:20.054612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T06:24:21.611234Z digest=sha256:1816a1991f84389663fb6e3d45cfc53373213df2c184cb53f5a3420d39d70dcd

Observation 3a8a1f20-b2c6-4787-8ae0-41dd02c5fc73 · inbound

Benchmarking Single-Factor Physical Video-to-Audio Generation cites this paper.

Benchmarking Single-Factor Physical Video-to-Audio Generation Video-to-Audio Generation with Hidden Alignment

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.377464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T07:41:56.917119Z digest=sha256:c7f29ab687257af3dda55247fa359b89d77b7393c36e21c36ffa076ec04a4335

Observation d377455d-877f-4951-8157-bcaeb1446c68 · inbound

Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer cites this paper.

Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer Video-to-Audio Generation with Hidden Alignment

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.296619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T21:17:48.886421Z digest=sha256:dd5fb1e1b96dbabba96bfe645b429017ca07f45ddb041c5fb8770ec36e45714d

Observation 74be9e54-8746-4131-b556-385d7d6d9475 · inbound

Foley-Omni: A Unified Multimodal Generation Model from Task-Level Audio Synthesis to Complete Video Soundtrack Generation cites this paper.

Foley-Omni: A Unified Multimodal Generation Model from Task-Level Audio Synthesis to Complete Video Soundtrack Generation Video-to-Audio Generation with Hidden Alignment

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T04:56:39.413881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T08:38:44.906160Z digest=sha256:6a3d137e0bc4538c26c898e8abb0432ce4cd81c543e6a204bcd408673e02b3de

Observation 66eed211-a3b8-45f9-88d3-94138eb509cb · inbound

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation cites this paper.

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation Video-to-Audio Generation with Hidden Alignment

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T14:02:31.121546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:02:31.121546Z digest=sha256:fe0c9499ce449e30516c5b147cb9fc08de4f79f76112cc08cb7064c37e4b82f7