Pith. sign in

Paper Citation Record · LEDGER

Temporally Aligned Audio for Video with Autoregression

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2409.13689.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.13689 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:43:41.137426Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T05:50:17.804576Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e6e8909a-634d-486e-bc02-f373299d398c · inbound

Video-Guided Foley Sound Generation with Multimodal Controls cites this paper.

Video-Guided Foley Sound Generation with Multimodal Controls Temporally Aligned Audio for Video with Autoregression

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T12:00:01.699988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:00:01.699988Z digest=sha256:27d63e5c954a5be801f8ac330295e1d604afc865fc8f838ca4fb2717f64c8e75

Observation 7af3061b-bd22-4409-9e03-c35538f734b8 · inbound

YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls cites this paper.

YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls Temporally Aligned Audio for Video with Autoregression

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T17:17:49.276216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:17:49.276216Z digest=sha256:40a0a65a72b1e4de4192d552d58da4e0edeeb0f3de712804183c39ec679a6a24

Observation 17c6a2ae-05ee-4646-b9bd-c3365752c2e5 · inbound

AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation cites this paper.

AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation Temporally Aligned Audio for Video with Autoregression

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T11:38:08.607273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:38:08.607273Z digest=sha256:03a0da14796a2b93ead94adeba9fdb7495c1b1d8eada5d8a8bfee4fec7d4a8f8

Observation 9bbad2fc-ca5d-4cb5-a382-3c2da3cdcd7f · inbound

MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis cites this paper.

MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis Temporally Aligned Audio for Video with Autoregression

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T11:35:57.834871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:35:57.834871Z digest=sha256:5cde8a48c039f0e6470a003f47f5cac45d9eb1463789a5c3aff5cf8ef1fec650

Observation 7ca429fd-5320-47ba-add0-d721791b7105 · inbound

Sound Scene Synthesis at the DCASE 2024 Challenge cites this paper.

Sound Scene Synthesis at the DCASE 2024 Challenge Temporally Aligned Audio for Video with Autoregression

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T20:27:03.788957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:27:03.788957Z digest=sha256:7c75998d1a3eec0aee461e7b58d92823c6b4a2f2ebffac910cf5fdfec9db2f37

Observation eaa969e7-3ab2-4901-a887-c4c5f3aef29c · inbound

OmniAudio: Generating Spatial Audio from 360-Degree Video cites this paper.

OmniAudio: Generating Spatial Audio from 360-Degree Video Temporally Aligned Audio for Video with Autoregression

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T11:43:41.137426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:43:41.137426Z digest=sha256:b6fcfd8a6f6d5ffb404845c0bcaa3f68594e59ffe588095fec7dcd66fe148f0b

Observation 8aab714f-f38e-46bd-abbc-3c3d4a717b46 · inbound

Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks cites this paper.

Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks Temporally Aligned Audio for Video with Autoregression

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:11.397850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:11.397850Z digest=sha256:3178ca17bd97e33ed98c6d72c91ac9fcbdf2bf1b05fbbfff50d9de70c762c8b4

Observation a879510a-756e-404b-aee7-f14b9d686219 · inbound

Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation cites this paper.

Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation Temporally Aligned Audio for Video with Autoregression

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:37:18.705798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:37:18.705798Z digest=sha256:b4b900d09764ac68b01d71e8b770abb81a2306c90a2f830e3ace39fe0a4a7660

Observation 702af5aa-9c8d-4f6e-8fa0-14ed86edf5ad · inbound

LD-LAudio-V1: Video-to-Long-Form-Audio Generation Extension with Dual Lightweight Adapters cites this paper.

LD-LAudio-V1: Video-to-Long-Form-Audio Generation Extension with Dual Lightweight Adapters Temporally Aligned Audio for Video with Autoregression

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T17:43:11.887020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:43:11.887020Z digest=sha256:9e639bc377632ff21a39c6708ba278f7d4886dfce1134144a72543330829e8ff

Observation 0fdf94d3-d15b-4f2c-90d6-0d5ded89b42b · inbound

Efficient Video-to-Audio Generation via Multiple Foundation Models Mapper cites this paper.

Efficient Video-to-Audio Generation via Multiple Foundation Models Mapper Temporally Aligned Audio for Video with Autoregression

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T05:50:17.811973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:50:17.701833Z digest=sha256:da03796f3cee895a4e995c1e85470dc4678721a08518d512a58d8fe6f9e4d1d4