Pith. sign in

Paper Citation Record · LEDGER

FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2407.01494.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.01494 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T05:50:17.728841Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T07:14:45.319897Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 74f1ff39-134b-41a1-b0fa-820359ac3e42 · inbound

Efficient Video-to-Audio Generation via Multiple Foundation Models Mapper cites this paper.

Efficient Video-to-Audio Generation via Multiple Foundation Models Mapper FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T05:50:17.728841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:50:17.728841Z digest=sha256:f7d063b935b77d84656bb34908e63219b78ebffeb343db67d3fb342aace5f4e2

Observation 0dcbd26a-08fd-4884-adeb-b9529976d844 · inbound

MeanFlow-Accelerated Multimodal Video-to-Audio Synthesis via One-Step Generation cites this paper.

MeanFlow-Accelerated Multimodal Video-to-Audio Synthesis via One-Step Generation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T23:46:32.481088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:46:32.481088Z digest=sha256:8c0be1ff0a8a28fdabf85d236cbcc717c67c34b30f5111aa0ebc4469be81a0c4

Observation 30b453a0-6a9a-4f6e-a2d1-f92dbce83eef · inbound

StereoFoley: Object-Aware Stereo Audio Generation from Video cites this paper.

StereoFoley: Object-Aware Stereo Audio Generation from Video FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:26:28.615296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T14:23:08.144599Z digest=sha256:885a05735f1225e026e2b0fb3005e1b7e37acea5b4521d53f344e4c64f080c66

Observation bb145c55-3032-4ca0-b7ad-a7d9cd1ec4c1 · inbound

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance cites this paper.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.863177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:5a809eaeee4deaa384f8671eb4de9a03fa1f3bf9b5dabf266be1f3eb99663e23

Observation e9523196-8c24-4383-8b41-a806591b514e · inbound

NoiseShift: Resolution-Aware Noise Recalibration for Better Low-Resolution Image Generation cites this paper.

NoiseShift: Resolution-Aware Noise Recalibration for Better Low-Resolution Image Generation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:44:22.914024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T21:41:29.502472Z digest=sha256:0a16bac0d26331d6c1db6f5ea6110c64bc7ca861902d549214a670e23267730d

Observation efdf7318-8259-4684-9be2-2691000d5af7 · inbound

MMAudioSep: Taming Video-to-Audio Generative Model Towards Video/Text-Queried Sound Separation cites this paper.

MMAudioSep: Taming Video-to-Audio Generative Model Towards Video/Text-Queried Sound Separation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:21:06.832653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T08:20:02.986562Z digest=sha256:df533ef64340ca23522d3d0acf42e6100e376215c308f2b9f6e17ed424a86f61

Observation 270cdb8c-cd4a-449e-b745-efa3bffbc071 · inbound

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection cites this paper.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.459651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:c433d2a04ce254596bda32ea9e154f226e21d78c9ba9cd748232b6f78bda9ffb

Observation fd5b86f8-7e4b-42c5-aa90-1aaa05de5c4d · inbound

JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing cites this paper.

JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T16:26:00.172036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:26:00.172036Z digest=sha256:823e0d588f27fa7760f07c4f75e5d759d75831948dc7a84e15b0d7b15403b4f2

Observation 38dca50c-d3e2-4ae5-8c0a-987523629705 · inbound

Aliasing-Free Neural Audio Synthesis cites this paper.

Aliasing-Free Neural Audio Synthesis FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:24.819824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:34:50.539351Z digest=sha256:f336b11ebddd752956e2592d5575686eb19e110cd93669762f7f9bdbd275078d

Observation aacf3a6c-c988-4296-baaf-0a0eb3aae606 · inbound

Aliasing-Free Neural Audio Synthesis cites this paper.

Aliasing-Free Neural Audio Synthesis FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T14:33:13.284553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:33:13.284553Z digest=sha256:5919fc4be8c8c8739c3efdc6e4b6b36bab10aa40b11bd73787eac0b68f2bc972

Observation 60b34967-c7dc-44ea-8326-832959ce304e · inbound

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation cites this paper.

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:48:21.873213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T19:43:37.604351Z digest=sha256:bae764dcf53706685d0770c7d3988f23afa8394fe5aaa5153827ddc40f0df6d5

Observation d368668f-7ac6-4ef1-af97-4dfc74e2fa66 · inbound

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation cites this paper.

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-21T16:14:15.286466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T16:10:31.015783Z digest=sha256:90d596e9d72623316e37eef229459e9302a0530c761013dcfd87e3743b00b5ad

Observation 024b790b-0065-4ec2-a200-819594430712 · inbound

EchoFoley: Event-Centric Hierarchical Control for Video Grounded Creative Sound Generation cites this paper.

EchoFoley: Event-Centric Hierarchical Control for Video Grounded Creative Sound Generation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T13:21:11.469829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:21:11.469829Z digest=sha256:88ffd7213582e27af1d830820e499a4c32364f7898ae6c5344a36f6833206e59

Observation 7c68a5e7-5887-4578-a2f9-067e55854cac · inbound

LTX-2: Efficient Joint Audio-Visual Foundation Model cites this paper.

LTX-2: Efficient Joint Audio-Visual Foundation Model FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:06:20.588495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:06:20.470686Z digest=sha256:72657f8d2ed8ff9acdbddaae7c849f2c66155ef4f4e2e4b411d8b645df715b27

Observation 7ba0f5e0-f392-44e4-9998-45b1d7a2da2e · inbound

Echoes Over Time: Unlocking Length Generalization in Video-to-Audio Generation Models cites this paper.

Echoes Over Time: Unlocking Length Generalization in Video-to-Audio Generation Models FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:56:33.536685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T19:53:18.200223Z digest=sha256:73ffcdda7d57f8571ab8d3c443c866df777bc9bafd042913d4d9987f344fc333

Observation 652fc853-fccd-41c3-a7cc-f63e44889002 · inbound

OmniSonic: Towards Universal and Holistic Audio Generation from Video and Text cites this paper.

OmniSonic: Towards Universal and Holistic Audio Generation from Video and Text FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:00:47.523966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T20:23:36.774359Z digest=sha256:24972730a4c79cfbf0880c171676baf21de79644491e05f6eb9f389bacf5b230

Observation dc06ec0a-a6f5-48ab-8001-6162eefc0fe6 · inbound

FoleyDesigner: Immersive Stereo Foley Generation with Precise Spatio-Temporal Alignment for Film Clips cites this paper.

FoleyDesigner: Immersive Stereo Foley Generation with Precise Spatio-Temporal Alignment for Film Clips FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:40:51.786766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:57:21.434793Z digest=sha256:d2cf71ec1d1dc6e8cc822f377f0ee287541bf6394704ff9bb6de85578713ff71

Observation 586985e3-390f-4ee7-8c32-fdf4ec9cb9eb · inbound

Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery cites this paper.

Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:59:03.509778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T09:52:35.741400Z digest=sha256:203ab346652fef93d4fcda8c992995958fe277760913a38609b85cbbc915bb42

Observation db1c1ca5-c3f1-4f33-9b6d-21f3e19ff80a · inbound

MMAudio-LABEL: Audio Event Labeling via Audio Generation for Silent Video cites this paper.

MMAudio-LABEL: Audio Event Labeling via Audio Generation for Silent Video FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:01:22.064276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T18:53:07.364051Z digest=sha256:d12b6e5f0ca17d424c859b08ee55131da516cc9eb7b106452097ac446e3235a2

Observation c78bf827-1d7e-43b2-9537-4b9a10808678 · inbound

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV cites this paper.

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:00.751695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:52:38.330851Z digest=sha256:f425d3ab8a752d6a8aaecd011f0102c996473eb2c5e2ed0983c3ebe6ab80daba

Observation ad1a69ed-29b4-4851-9192-a1e478a4ffd5 · inbound

Benchmarking Single-Factor Physical Video-to-Audio Generation cites this paper.

Benchmarking Single-Factor Physical Video-to-Audio Generation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.374668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T07:41:56.917119Z digest=sha256:2d5325eb503018b660b51bf206bbaed3464e771ea63ef2c15331df5c78e34330

Observation b36f4850-85bc-4d41-ac98-0c0103d423e0 · inbound

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation cites this paper.

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:28:18.788790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T08:04:48.283908Z digest=sha256:228f20fe0e0e4468f89ecd5d32f09da0f2dd01d21bb7b2290f8024f43048c732

Observation 25a17be2-65fb-4005-b514-56a14b719d22 · inbound

Precise Video-to-Audio Generation with Cross-Modal Alignment in Latent Space cites this paper.

Precise Video-to-Audio Generation with Cross-Modal Alignment in Latent Space FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-08T07:14:45.321352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-08T07:10:20.519266Z digest=sha256:eb3e7df5f0a58070bc37559ca741b0492cee4899b1300648e1277b0b79919db7

Observation bdea5d69-94cc-4cd0-a427-0d4cbc8e5e53 · inbound

Precise Video-to-Audio Generation with Cross-Modal Alignment in Latent Space cites this paper.

Precise Video-to-Audio Generation with Cross-Modal Alignment in Latent Space FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T08:23:35.864870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:23:35.864870Z digest=sha256:a013affc87969b132d7d0e0648c3e881579a9f6680cf674a446abd7af7d862ea