Pith. sign in

Paper Citation Record · LEDGER

Generative AI for Music and Audio

As of 12 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2411.14627.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14627 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:10:14.185025Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c857bac9-c5f9-4aee-a014-c1460a8aa07d · outbound

This paper cites TensorFlow: A system for large- scale machine learning.

Generative AI for Music and Audio TensorFlow: A system for large- scale machine learning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.286039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:10:14.148529Z digest=sha256:4fe3c07709313db16f36c98dedd77be64f7452cd787febad346213b33e8abe20

Observation 7c08aaaf-e2c3-4480-aca2-dcde314eee68 · outbound

This paper cites Noise2Music: Text-conditioned Music Generation with Diffusion Models.

Generative AI for Music and Audio Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.151790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.151790Z digest=sha256:87d0699623468827d3d08eb770b2c6732cfa3cfa335c1dc4dd246cfc7d639e6a

Observation e7d6b493-1a91-4c74-a6a9-089b179f23bf · outbound

This paper cites BERT-like Pre-training for Symbolic Piano Music Classification Tasks.

Generative AI for Music and Audio BERT-like Pre-training for Symbolic Piano Music Classification Tasks

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.155450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.155450Z digest=sha256:bafcb6fa5e29a8b3db226cd183136cae01fe04bfd7a25cc20f2aa76a8385d613

Observation db998460-ee34-4d73-8e1b-6c7ca5feccfa · outbound

This paper cites MMM : Exploring Conditional Multi-Track Music Generation with the Transformer.

Generative AI for Music and Audio MMM : Exploring Conditional Multi-Track Music Generation with the Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.158686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.158686Z digest=sha256:581aefe868d85af075b9ce111e6b168f03ad0f21041be52a120209110bf0e4dd

Observation 340f063d-6c05-4ffb-b852-b180a4adcf97 · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

Generative AI for Music and Audio AudioGen: Textually Guided Audio Generation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.278188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:10:14.161825Z digest=sha256:6761b584ee06abb50385aa5d42cf0e8a4d00cd36b054ed8b72272e5610752deb

Observation 5c4b5c12-52d4-4b05-908f-d3cd3198861b · outbound

This paper cites WavJourney: Compositional Audio Creation with Large Language Models.

Generative AI for Music and Audio WavJourney: Compositional Audio Creation with Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.164467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.164467Z digest=sha256:24a8460df858783548ca1cf613c10fe7fb6686a10d4efb38701ce8ffe495a2c7

Observation eeeb6c09-e0c6-4fa0-bd46-be027c669473 · outbound

This paper cites PyTorch: An Imperative Style, High-Performance Deep Learning Library.

Generative AI for Music and Audio PyTorch: An Imperative Style, High-Performance Deep Learning Library

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.271008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:10:14.167567Z digest=sha256:48578b5b3bd9e5a7b4e5ffd01fc5827bfd7da240e86d59314682a2333ee1292a

Observation 27d1484d-0f59-4b0f-aabe-58c6b8808f32 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Generative AI for Music and Audio Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.170530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.170530Z digest=sha256:9ba519dd067d04e8e76b4897d0c0362b53ba2cb4f083a840658c43745b05838b

Observation ecc9a73f-d6cc-4680-8026-488964378b0c · outbound

This paper cites DeepSinger: Singing Voice Synthesis with Data Mined From the Web.

Generative AI for Music and Audio DeepSinger: Singing Voice Synthesis with Data Mined From the Web

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.263856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:10:14.173330Z digest=sha256:c36088ef1ae67f6e26923928f03adcfcdb81d161fe8066b92b0e554a8508cd89

Observation d5362403-a65a-4566-9025-37955489269a · outbound

This paper cites FIGARO: Generating Symbolic Music with Fine-Grained Artistic Control.

Generative AI for Music and Audio FIGARO: Generating Symbolic Music with Fine-Grained Artistic Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.176396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.176396Z digest=sha256:81f16a2cd3b68cdc79802e46e74d830cee87f9c3373b2228010e26beb07c70b0

Observation ec8ed454-0267-4a88-b1af-d4c5fffb268a · outbound

This paper cites LAION-5B: An open large-scale dataset for training next generation image-text models.

Generative AI for Music and Audio LAION-5B: An open large-scale dataset for training next generation image-text models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.254659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:10:14.179278Z digest=sha256:3e80b5296b53b584869f6b8d4cb740cec2468bd4ee930f05ccc284f47bf56e71

Observation b32daa7e-6e65-46e0-82b4-2ba872cd9d36 · outbound

This paper cites A Survey on Neural Speech Synthesis.

Generative AI for Music and Audio A Survey on Neural Speech Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.182250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.182250Z digest=sha256:6fac6264852c1e1b4d6e44c4bd8f16daa93ae37a311bc0c3febc3ab949fe1b1e

Observation d27ae42c-27ac-42ed-ba55-fc6a84f16ab6 · outbound

This paper cites Diffsound: Discrete Diffusion Model for Text-to-sound Generation.

Generative AI for Music and Audio Diffsound: Discrete Diffusion Model for Text-to-sound Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.185025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.185025Z digest=sha256:411e5c7f4328a9227dd76ee9397b407cc9d67c2a569f0a3b8e40558939aaff02

Pith citing papers

No inbound Pith citation observations are available.