Pith. sign in

Paper Citation Record · LEDGER

Long-form music generation with latent diffusion

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2404.10301.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.10301 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:59:46.716586Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T00:04:22.451108Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 351fe317-95a3-40ac-a656-48523bf94e57 · inbound

In-the-wild Audio Spatialization with Flexible Text-guided Localization cites this paper.

In-the-wild Audio Spatialization with Flexible Text-guided Localization Long-form music generation with latent diffusion

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:46.716586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:59:46.716586Z digest=sha256:d2b48acfafd178fced68c8931d8fad167a599617b9f9350e2a377cd471a98683

Observation 237bc281-e6e2-478d-99bc-f3e97e626b97 · inbound

WAKE: Watermarking Audio with Key Enrichment cites this paper.

WAKE: Watermarking Audio with Key Enrichment Long-form music generation with latent diffusion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T10:19:26.350553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:19:26.350553Z digest=sha256:47117d2c33d6e8dca5794dffb37071b482e7d57be9ce9662fc58216c00e145a1

Observation c272e73f-2a7a-487d-9aae-b8c315fdf398 · inbound

Music Boomerang: Reusing Diffusion Models for Data Augmentation and Audio Manipulation cites this paper.

Music Boomerang: Reusing Diffusion Models for Data Augmentation and Audio Manipulation Long-form music generation with latent diffusion

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:42:05.689145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:42:05.689145Z digest=sha256:6c2505c1ff87747fc25c8fcb7ea37cd38b1f1f86130ae3171e10e1b900682aba

Observation 10d9b764-8e6b-474f-819d-61db03b33a6b · inbound

ASAudio: A Survey of Advanced Spatial Audio Research cites this paper.

ASAudio: A Survey of Advanced Spatial Audio Research Long-form music generation with latent diffusion

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T22:54:54.920307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:54:54.920307Z digest=sha256:12931fe46121c051374dcd1cac4c3d8c4dca711e6c2d333fdc9c0ec4fde939fe

Observation d630949b-9041-49c4-b6a2-89e5cf734cab · inbound

Woosh: A Sound Effects Foundation Model cites this paper.

Woosh: A Sound Effects Foundation Model Long-form music generation with latent diffusion

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:15.857234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:51:08.144573Z digest=sha256:3ffed609512633ae5237bfe44c2ec4414f2b0a128a7f8c4414519c07464f53c5

Observation 1684b2a8-9e63-471a-bbc4-f3debbe0da5b · inbound

Seconds-Aligned PCA-DAC Latent Diffusion for Symbolic-to-Audio Drum Rendering cites this paper.

Seconds-Aligned PCA-DAC Latent Diffusion for Symbolic-to-Audio Drum Rendering Long-form music generation with latent diffusion

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T18:39:22.183244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T18:38:43.559539Z digest=sha256:5e2600fb75764390e73cfaacca9d9f53695946c497ef650f2515b73fc8111f09

Observation 180e0950-93d0-4cab-9479-d8fe8e82756b · inbound

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators cites this paper.

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators Long-form music generation with latent diffusion

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-22T03:25:58.995820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T03:24:50.604019Z digest=sha256:561b42f3d94c3abe171c12f6449ef08e836a34ca6fbfcaffe34bddd478f5240a

Observation 71065932-b94a-463a-8a0a-5fdda690985a · inbound

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation cites this paper.

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation Long-form music generation with latent diffusion

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:28:18.779223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T08:04:48.283908Z digest=sha256:e1336b173d380db470cf01de11e5f83424f0bf5162997ea5d6206114a52e5c3d

Observation 1d69060d-9e2a-4a32-a8e5-3b446e87df76 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence Long-form music generation with latent diffusion

Reference 111

Resolution
verified exact
local_arxiv, observed 2026-07-08T00:04:22.452391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:1d4742f60567a5dbf961cf45ac1ad0796ef67c4a67648be09e27fd514d0d27ad

Observation d6a3e05c-0ad7-42ab-9d30-eb6c2525024d · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence Long-form music generation with latent diffusion

Reference 111

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:896492a832c600ffeb77db9c7dae29cd75672fba5668f8dd81bd2ca5a61987cf

Observation 85dbf6a9-efbb-4786-8349-a07511050edd · inbound

Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens cites this paper.

Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens Long-form music generation with latent diffusion

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T08:35:48.723618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:35:48.723618Z digest=sha256:65bee216a05e1360a1b1b2143cda6d073fca836c712d03d3980925ec9f823161

Observation be5684af-b5ae-4839-a628-2b5371d2a9eb · inbound

On the Geometry of Music Bandwidth Extension in Latent Spaces of Audio Codecs cites this paper.

On the Geometry of Music Bandwidth Extension in Latent Spaces of Audio Codecs Long-form music generation with latent diffusion

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T14:08:31.796045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:08:31.796045Z digest=sha256:a194db864b628318b8007874fb2c7824ae4dea5fcfb46fad5a83eaacd6b63928

Observation 775a07d5-4098-423c-b0a6-ba1b3b538eeb · inbound

AI-Based Sound Effect Generation: A Narrative Review of Generative Models Across Input Modalities cites this paper.

AI-Based Sound Effect Generation: A Narrative Review of Generative Models Across Input Modalities Long-form music generation with latent diffusion

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T13:43:18.097916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:43:18.097916Z digest=sha256:6cbf27fbaab2bcc9ac4f8e0fcaed5ac4faaf6dd9b0c5ec40b5fe84964bbcfd92