Pith. sign in

Paper Citation Record · LEDGER

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation

As of 14 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2607.29148.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.29148 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T12:47:00.255183Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b7faae89-7517-47d2-8fbd-566671c26b64 · outbound

This paper cites Foleygan: Visu- ally guided generative adversarial network-based syn- chronous sound generation in silent videos,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Foleygan: Visu- ally guided generative adversarial network-based syn- chronous sound generation in silent videos,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.420921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.420921Z digest=sha256:b11741a6391cb43c5f65eb2ec248afbb6a6d6b09c39fc44461c658f087e4c5af

Observation e2e21079-58d3-4071-840c-8f226764a765 · outbound

This paper cites The x-lance system for dcase2023 challenge task 7: Foley sound synthesis track b,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation The x-lance system for dcase2023 challenge task 7: Foley sound synthesis track b,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.471416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.471416Z digest=sha256:0ed2b89a565ff87c4af7543cec9df326fb58ddd43ee26865c49935bd78517bc5

Observation 44e02b7d-27d6-494a-adef-748b392c51f2 · outbound

This paper cites Mtdiffusion: Multi- task diffusion model with dual-unet for foley sound generation,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Mtdiffusion: Multi- task diffusion model with dual-unet for foley sound generation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.566213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.566213Z digest=sha256:819aaabcb6a35fc2804abb824f85942056a3dbecaf49f3c21cd1b4908b579436

Observation a8cc5034-618d-477f-bfa9-f77711cedb6a · outbound

This paper cites Rhythmic foley: A framework for seamless audio-visual alignment in video-to-audio synthesis,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Rhythmic foley: A framework for seamless audio-visual alignment in video-to-audio synthesis,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.673463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.673463Z digest=sha256:fb3511ec09c4f235f65ae9a8852d8690a18a1a7d9fccc13f89b36df15001f66a

Observation 8b3ef8bd-5fbc-4dbc-8fc3-b9955bbdad09 · outbound

This paper cites T-foley: A controllable waveform-domain diffusion model for temporal-event- guided foley sound synthesis,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation T-foley: A controllable waveform-domain diffusion model for temporal-event- guided foley sound synthesis,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.781442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.781442Z digest=sha256:66cc10370473c66b745e178ae60490dec069993d00999cadc2091a26b9f7690f

Observation 66d2d469-9ce7-4885-ada3-0d6d8c8eec1f · outbound

This paper cites Text-Driven Foley Sound Generation With Latent Diffusion Model.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Text-Driven Foley Sound Generation With Latent Diffusion Model

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.862868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.862868Z digest=sha256:ddd212edbf6d291268e4bc148e54040c4d909c15cad37b757b086a6e1907a980

Observation 8e6509e6-dac5-48c8-9391-0c7876ebc211 · outbound

This paper cites AudioLDM: Text-to-Audio Generation with Latent Diffusion Models.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.947785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.947785Z digest=sha256:54e776e802663b2358381aff1062024f4cc75aef1cc9c21cb0c445ff4e1e9864

Observation 13aa4e11-f584-4fb5-a94b-3f1d056adb00 · outbound

This paper cites Audioldm 2: Learning holistic audio generation with self-supervised pretraining,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Audioldm 2: Learning holistic audio generation with self-supervised pretraining,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.077301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.077301Z digest=sha256:0bbb172daa310d3bd0f173505f3fa1afce208b9a59f6cb74e17c4657fd53ae7f

Observation 5fcd8b7b-cf3a-46e5-8093-3caeff0681c3 · outbound

This paper cites Generative speech foundation model pre- training for high-quality speech extraction and restora- tion,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Generative speech foundation model pre- training for high-quality speech extraction and restora- tion,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.169116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.169116Z digest=sha256:4fe2989ea0e9c0aa159032ff8d69416fb7ecddbe360e97ad91e68318709a341a

Observation fc5404f4-d96a-4b8e-8479-6aa933bd6c08 · outbound

This paper cites Dual-path rnn: Ef- ficient long sequence modeling for time-domain single- channel speech separation,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Dual-path rnn: Ef- ficient long sequence modeling for time-domain single- channel speech separation,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.293363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.293363Z digest=sha256:e0447fc2e888960ddc689f04cb800234d9bdc6d4e3370fd97f64db10970f3831

Observation 6346c9a8-642c-4cbc-870a-67acc081768c · outbound

This paper cites A study on speech enhancement based on diffusion probabilistic model,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation A study on speech enhancement based on diffusion probabilistic model,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.406840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.406840Z digest=sha256:0ee5b334e16d44c0dcb7e4d68c6c899c9b09fb90f117708ccad6ee6d604b3628

Observation f42dd0bd-c244-476c-b2fe-43aa44a00dcf · outbound

This paper cites Unsupervised single-channel audio sep- aration with diffusion source priors,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Unsupervised single-channel audio sep- aration with diffusion source priors,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.456317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.456317Z digest=sha256:4eceb8fb9aaac884a5ea8592ece631e2af66312e2f0f98fe2e61119a2f068758

Observation 367a5819-130d-4faa-9a77-92d11bd494ac · outbound

This paper cites Mambafoley: Foley sound generation using selective state-space models,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Mambafoley: Foley sound generation using selective state-space models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.525316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.525316Z digest=sha256:018397ce9508838b19879eb7b67e4a217bbb0e4fb52393776172773f9bc0d23a

Observation f6f8ba3a-c65d-476a-b286-830334482bc2 · outbound

This paper cites Full-band general audio synthesis with score- based diffusion,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Full-band general audio synthesis with score- based diffusion,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.615845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.615845Z digest=sha256:9a8d0035719b192bd10a05cd5260c9f6ca6d56b7882c2797f58266424a28b121

Observation cacf177a-1e36-4216-8978-fd6a2a216346 · outbound

This paper cites Diffwave: A versatile diffusion model for audio synthesis,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Diffwave: A versatile diffusion model for audio synthesis,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.663901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.663901Z digest=sha256:1a2268bbbb29528481d84bbdbb60efa7e8ef2dc867b93256a83ec2d8178197dc

Observation 14582748-2ac1-4ebd-9b42-f39e00e5c434 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.714092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.714092Z digest=sha256:b0f8a6df967677861ecd3b48d971aa269d342367f30326e942a6157cdb506e9a

Observation 56df86ed-dd3c-467a-a138-6a7803d005f7 · outbound

This paper cites Fregrad: Lightweight and fast frequency-aware diffusion vocoder,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Fregrad: Lightweight and fast frequency-aware diffusion vocoder,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.773065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.773065Z digest=sha256:0c8f4dad590de009c15ca608cb060154511494f68cb1a06909dd6d0bc6ccb15f

Observation 2eea032b-d09e-4b00-a289-3660bdc6e82c · outbound

This paper cites UnDiff: Unsupervised Voice Restoration with Unconditional Diffusion Model.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation UnDiff: Unsupervised Voice Restoration with Unconditional Diffusion Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.820387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.820387Z digest=sha256:b20678a121821f02235f530c8d5a2e42bb6cf5a820154fcfbd07f9a9dbc19a52

Observation 690cc6c0-2e86-4641-8332-34544f01c746 · outbound

This paper cites Solving audio inverse problems with a diffusion model,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Solving audio inverse problems with a diffusion model,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.906532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.906532Z digest=sha256:93f6b57ab73a390979dec7d9c9fb742dd26fb6f8f1165fb2958133dd647958df

Observation 65e250d6-84ca-4ec1-b5c3-87298b4826ae · outbound

This paper cites Foley Sound Synthesis at the DCASE 2023 Challenge.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Foley Sound Synthesis at the DCASE 2023 Challenge

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.967862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.967862Z digest=sha256:afb4a91a40043d2aa68c6912afec665424476671d45c19ddde46d6a8185c28ee

Observation c9b1bd9d-29cc-4567-9f64-681b9d0ba6b2 · outbound

This paper cites Scalable diffusion models with transformers,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Scalable diffusion models with transformers,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T12:47:00.032718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:47:00.032718Z digest=sha256:1b8552132e2514515f9b78c0838545c08f12b6ba4b7008a2acaf5eca131e4cf3

Observation b3d4713e-705c-430f-abac-c7374293af9e · outbound

This paper cites General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T12:47:00.111103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:47:00.111103Z digest=sha256:949fce24dceb6ff4d452c81c8e724aa829d64ec1cc1d4d8ca5046e7763c2141a

Observation cf6d8bac-7564-4a98-8948-0629a54414e6 · outbound

This paper cites Decoupled Weight Decay Regularization.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Decoupled Weight Decay Regularization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T12:47:00.165216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:47:00.165216Z digest=sha256:27e8a8976447871abfcf231499de80bb86a30cd9a1af4196238949b1f70c77d5

Observation b26af08f-57c4-46ea-8b12-a2f75bbfd77a · outbound

This paper cites Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T12:47:00.255183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:47:00.255183Z digest=sha256:5e42959fb78f51e875461fffeac1ff26873124114dcb717de464f2711f7638c4

Pith citing papers

No inbound Pith citation observations are available.