Pith. sign in

Paper Citation Record · LEDGER

Noise2Music: Text-conditioned Music Generation with Diffusion Models

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 46 inbound Pith citation observations for arXiv:2302.03917.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.03917 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 46 of 46 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:11:21.714787Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T00:04:22.608820Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2aac7385-01d9-46ac-9a52-5ffe35473201 · inbound

Shap-E: Generating Conditional 3D Implicit Functions cites this paper.

Shap-E: Generating Conditional 3D Implicit Functions Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T15:32:06.770084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T15:32:06.563955Z digest=sha256:c756519234de896c80884a8196794297f04471618a8b8104987e141cf430415c

Observation eca944cf-56c8-4bf3-8dff-1ab0605dd869 · inbound

Movie Gen: A Cast of Media Foundation Models cites this paper.

Movie Gen: A Cast of Media Foundation Models Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:23.975684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T14:16:18.521699Z digest=sha256:d294ed489c737d59c0e3bac1763ab0a985963c037f4f4b1b8798102ea6d9abfd

Observation 16bed474-6abc-4052-9914-27fdb1e4c85f · inbound

Instruction-Guided Editing Controls for Images and Multimedia: A Survey in LLM era cites this paper.

Instruction-Guided Editing Controls for Images and Multimedia: A Survey in LLM era Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-12T20:10:24.442867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:10:24.442867Z digest=sha256:96d756a1e65d2552f3c7711d49a7b204fe5a03e307f261ae61cc0e266d06d405

Observation 7c08aaaf-e2c3-4480-aca2-dcde314eee68 · inbound

Generative AI for Music and Audio cites this paper.

Generative AI for Music and Audio Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.151790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.151790Z digest=sha256:d120adeade132a86de2fbfea5b7d6dc01365cbbefeadab3c656f74fcd7c96ff2

Observation 4d2759e9-e5bf-4ab9-9810-e44e547a7987 · inbound

Repurposing Image Diffusion Models for Training-Free Music Style Transfer on Mel-spectrograms cites this paper.

Repurposing Image Diffusion Models for Training-Free Music Style Transfer on Mel-spectrograms Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:35:29.732673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-23T08:32:51.237590Z digest=sha256:6b9e3693e92c07e4766ce0cb472237b99ae8e24055d38d1fe36425eb3b637e15

Observation 8983eb6c-79e9-4b53-b801-0bdc250b38fa · inbound

DiffSLT: Enhancing Diversity in Sign Language Translation via Diffusion Model cites this paper.

DiffSLT: Enhancing Diversity in Sign Language Translation via Diffusion Model Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T12:24:55.268545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:24:55.268545Z digest=sha256:d52b218dc97947ca2eb22432f0bb97ca277bb97ed1f9e9ec1e9477dc96f54939

Observation b4715119-363c-4ec9-a92f-db576b4fda12 · inbound

VidMusician: Video-to-Music Generation with Semantic-Rhythmic Alignment via Hierarchical Visual Features cites this paper.

VidMusician: Video-to-Music Generation with Semantic-Rhythmic Alignment via Hierarchical Visual Features Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:26.999000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:26.999000Z digest=sha256:da859b10467bf2dc809840a485ed50e27e433da256b6f449a86b8106b1cf191c

Observation 10b10c40-ff12-4466-b00e-5aa37351cb84 · inbound

Does it Chug? Towards a Data-Driven Understanding of Guitar Tone Description cites this paper.

Does it Chug? Towards a Data-Driven Understanding of Guitar Tone Description Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:04.221749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:04.221749Z digest=sha256:2ebf7c030c2051de89b0e90ec1d4a7236491dcc528edd96e28a80eb5074dbf55

Observation 18b89a4a-0dd4-4166-966e-9f3462845c15 · inbound

Text2midi: Generating Symbolic Music from Captions cites this paper.

Text2midi: Generating Symbolic Music from Captions Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T10:32:45.997984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:32:45.997984Z digest=sha256:3e6a15ed7db781807975adbeb2fc1349f6eadc20745981b0d6edcef61a09d92b

Observation 5cac7aad-e1ee-42cf-af11-027aca01e085 · inbound

Amuse: Human-AI Collaborative Songwriting with Multimodal Inspirations cites this paper.

Amuse: Human-AI Collaborative Songwriting with Multimodal Inspirations Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:20.654142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:23:20.654142Z digest=sha256:fb151fda234269fc5bbbd86f78fd232d64ed302a53f283153fcc264c237fa9ce

Observation 1ae4b6d1-0c62-45de-b91b-a1aa2d975d35 · inbound

ETTA: Elucidating the Design Space of Text-to-Audio Models cites this paper.

ETTA: Elucidating the Design Space of Text-to-Audio Models Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T00:45:19.755042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:45:19.755042Z digest=sha256:b3be2d3155cf611ffb38965613ff4875b179e1658c119058a04bc89b02a3eeca

Observation 0e12e24b-30a3-4e14-a5c2-58df61beba44 · inbound

Tri-Ergon: Fine-grained Video-to-Audio Generation with Multi-modal Conditions and LUFS Control cites this paper.

Tri-Ergon: Fine-grained Video-to-Audio Generation with Multi-modal Conditions and LUFS Control Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T23:28:05.243917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:28:05.243917Z digest=sha256:1d3812c5764648ea934de44603958e55c7636e1cb27f3e7131c8057177c84be5

Observation 073131ab-8902-4418-92dc-9dbe19747a8e · inbound

XMusic: Towards a Generalized and Controllable Symbolic Music Generation Framework cites this paper.

XMusic: Towards a Generalized and Controllable Symbolic Music Generation Framework Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T20:22:28.080677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:22:28.080677Z digest=sha256:80819897c44d3e3595b8af72313a3c621f39dc1941838151a218007703875a60

Observation 595b6822-e9fd-40f2-b3de-95f4545bcdda · inbound

Latent Swap Joint Diffusion for 2D Long-Form Latent Generation cites this paper.

Latent Swap Joint Diffusion for 2D Long-Form Latent Generation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T20:12:29.458700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:12:29.458700Z digest=sha256:f2125021e60a6de92f9591a4dfb2f8ed6315a280f7bdd45fbcf015a7dc22b84d

Observation 0ced350e-1c9f-4f23-a5b8-0f6fa6c6fc80 · inbound

Music for All: Representational Bias and Cross-Cultural Adaptability of Music Generation Models cites this paper.

Music for All: Representational Bias and Cross-Cultural Adaptability of Music Generation Models Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T13:09:46.440827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:09:46.440827Z digest=sha256:d3a450252ca71910dda043c986057874966695f7cf549256925e95894b42ba76

Observation 72d3c04e-dfc8-4e57-b219-88712ec7df77 · inbound

JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata cites this paper.

JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T12:43:30.119680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:43:30.119680Z digest=sha256:5544e7800692ec7872834b85a9ce08cebfb8cfd09a6d9bb8a526be4ef7290b00

Observation a56ba5e3-1794-4899-9743-b7378d6eac04 · inbound

YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation cites this paper.

YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T05:10:06.205335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:10:06.205335Z digest=sha256:b8aa0a59839430b90381d2e8d9d6ed6c06f9b26d5b50a320b6380aba21b6c7dc

Observation d9ce4144-7390-49cf-9355-917ad3029a93 · inbound

MusFlow: Multimodal Music Generation via Conditional Flow Matching cites this paper.

MusFlow: Multimodal Music Generation via Conditional Flow Matching Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T12:11:21.714787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:11:21.714787Z digest=sha256:7985d4e3c4e6f508f55f778222aac7448fdd42fcabe4b032b4f2ae2bdcda587c

Observation 34d002d5-98ce-4316-a295-4033acfe07c6 · inbound

SongEval: A Benchmark Dataset for Song Aesthetics Evaluation cites this paper.

SongEval: A Benchmark Dataset for Song Aesthetics Evaluation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:38.877554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:38.877554Z digest=sha256:673844512dc9f44be0b7d11d05d9da4655aa09c89f52ee27df637a8a4a3ea412

Observation 2fad8cd0-e7bb-45f2-9f53-648d33483d67 · inbound

Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment cites this paper.

Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:34:08.660870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:34:08.660870Z digest=sha256:b996724a30c24b6bb58610cabc3fd1a3ae8c86129db717d2f3bec7dfdbbf6354

Observation 2e992f79-4caa-4c9e-aa50-0cc9b025c4c8 · inbound

Auto-Regressive vs Flow-Matching: a Comparative Study of Modeling Paradigms for Text-to-Music Generation cites this paper.

Auto-Regressive vs Flow-Matching: a Comparative Study of Modeling Paradigms for Text-to-Music Generation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T05:12:24.656553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:12:24.656553Z digest=sha256:f027e147d8eaa7242a5aaa4e3525f4e5bfecbcd85278566e92bf55770b388e59

Observation 92b97fe1-1c0a-456e-a812-f804e48e40f9 · inbound

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections cites this paper.

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:49:43.230631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:49:43.230631Z digest=sha256:7d591f84eff679cf4f1d264cb35a9d20988611cd1e2b6d2dc44153a607eb2257

Observation 5f4478ee-a3e7-4058-8c70-4eea61bc3156 · inbound

Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models cites this paper.

Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:58:33.511112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:58:33.511112Z digest=sha256:8b772d828fbecdac3f7fa4c73a33abc3ae2cfb603b5bed2b3bfc7d638b3708f9

Observation 7a6206e4-2e84-4a1a-ae3a-370afd45f416 · inbound

DefFusionNet: Learning Multimodal Goal Shapes for Deformable Object Manipulation via a Diffusion-based Probabilistic Model cites this paper.

DefFusionNet: Learning Multimodal Goal Shapes for Deformable Object Manipulation via a Diffusion-based Probabilistic Model Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:33.008408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:33.008408Z digest=sha256:4ba5dd13c8162cdff37eca69d7d447a3b1e3d73e9679f81e750ed78e609865c4

Observation 948da723-3b9f-4bab-9a97-4ebdfff51363 · inbound

Benchmarking Music Generation Models and Metrics via Human Preference Studies cites this paper.

Benchmarking Music Generation Models and Metrics via Human Preference Studies Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:41:08.296944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:41:08.296944Z digest=sha256:89e0de834349feea84920944ca3b97cdd5970d27e4119789b30f70444c533f08

Observation 0313e234-525e-4f9c-a8f1-f82da230d115 · inbound

Guidance in the Frequency Domain Enables High-Fidelity Sampling at Low CFG Scales cites this paper.

Guidance in the Frequency Domain Enables High-Fidelity Sampling at Low CFG Scales Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T18:31:43.607320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:31:43.607320Z digest=sha256:ada9d57c4e4a1bd77842b0a0fd36c7d7b02fa7ac3e4d8a08169a1653b4b73822

Observation c0f3978c-ab5d-443c-bfe1-9135a99d1740 · inbound

MusGO: A Community-Driven Framework For Assessing Openness in Music-Generative AI cites this paper.

MusGO: A Community-Driven Framework For Assessing Openness in Music-Generative AI Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T20:09:33.815726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:09:33.815726Z digest=sha256:a06b9f2e3ef3cf96c022e70dad60ef6bb9967942dad0915c71e47088e7138dd5

Observation edeebd84-0513-45ed-bc21-0844fbe762e9 · inbound

Hear-Your-Click: Interactive Object-Specific Video-to-Audio Generation cites this paper.

Hear-Your-Click: Interactive Object-Specific Video-to-Audio Generation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:39:54.151850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:39:54.151850Z digest=sha256:01fcac6c6e36ae76637153c1f4f932cf2c2ed2ed01feb5ab03993a9d7cf9898d

Observation 578d84d0-bd93-44ac-8b77-23fdcd9af446 · inbound

ASTAR-NTU solution to AudioMOS Challenge 2025 Track1 cites this paper.

ASTAR-NTU solution to AudioMOS Challenge 2025 Track1 Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T17:48:05.438833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:48:05.438833Z digest=sha256:0053689dfc598db068726654bf6974ae80ee28b90d23fabfa69fa66370b873b8

Observation 6d7c490a-eda3-4d2b-93c4-5e3ddbb3a816 · inbound

DiffRhythm+: Controllable and Flexible Full-Length Song Generation with Preference Optimization cites this paper.

DiffRhythm+: Controllable and Flexible Full-Length Song Generation with Preference Optimization Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:40:49.218862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:40:49.218862Z digest=sha256:ef66ec236c5250de54ae11acb3c415ce6316f95f0c5f05b9397cdc50097714f6

Observation 835621e8-4d12-4349-af34-4f93ca4f668a · inbound

Latent Fourier Transform cites this paper.

Latent Fourier Transform Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:21:07.015691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T03:45:07.892234Z digest=sha256:76ae47a7494df2a00fc53901466f0edbb49ef8241feb71cab3bab3f0e8cb4daf

Observation 3b8d8a75-84f4-4d56-8c02-060757903951 · inbound

S2Accompanist: A Semantic-Aware and Structure-Guided Diffusion Model for Music Accompaniment Generation cites this paper.

S2Accompanist: A Semantic-Aware and Structure-Guided Diffusion Model for Music Accompaniment Generation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-19T22:52:50.509604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T22:50:08.321514Z digest=sha256:74875d4f7e19d688f2b1b6e9a67789bbfafd0ca83856675e43b37352b00dc9ad

Observation 6703c3df-9d62-4e89-bec4-759e24a71161 · inbound

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions cites this paper.

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T00:56:24.705379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T13:10:29.213917Z digest=sha256:f734d5143201bad48415eb48f92eb04c957544fb30ff84d87509943148fafe68

Observation ad8a4e55-1445-49d6-8647-b606405e5c84 · inbound

Real-Time Interactive Music Generation via Data-Free Streaming Consistency Distillation cites this paper.

Real-Time Interactive Music Generation via Data-Free Streaming Consistency Distillation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:40:02.633895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-25T22:42:57.448084Z digest=sha256:8af168b63c112513a68458b4661bce0d8fad22e7a7c664bf1bb31bfd2bd2bc9f

Observation 9db21f8c-1fc6-4015-a917-b6380c2c3b29 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 86

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.610666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:4f4cb74e213555ac9f50014219c7c697866014fd4d66eb02491ce30a7b79425a

Observation 8418f976-53c8-4112-8b00-ae6de539a584 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:7cb69f9979675075e17cf3c92e7a75fbc5b0c0a6f175175319e54c49f3138530

Observation 98879aa0-8241-4e01-830b-39094736f6fd · inbound

Dance to Music Generation leveraging Pre-training with Unpaired data and Contrastive Alignment cites this paper.

Dance to Music Generation leveraging Pre-training with Unpaired data and Contrastive Alignment Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T10:59:22.914974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:59:22.914974Z digest=sha256:40663c66220548cf0b21b3349df06ab08c895717552a8af25318ee7e43e3792b

Observation 3683c75c-7bf9-41af-b7d9-61610cab0cec · inbound

Anysynth:Zero-Shot Instrument Cloning via In-Context Learning and Asymmetric Hierarchical Guidance cites this paper.

Anysynth:Zero-Shot Instrument Cloning via In-Context Learning and Asymmetric Hierarchical Guidance Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T06:45:43.330341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:45:43.330341Z digest=sha256:0815e862a384d462e56b5cb12f2fb0b425c18bd77799f602ff019a0d72706e4e

Observation 2dc6e734-ab82-41b3-b91a-18b412ca60aa · inbound

Anysynth:Zero-Shot Instrument Cloning via In-Context Learning and Asymmetric Hierarchical Guidance cites this paper.

Anysynth:Zero-Shot Instrument Cloning via In-Context Learning and Asymmetric Hierarchical Guidance Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T07:05:04.187324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:05:04.187324Z digest=sha256:096767b9a7ac7fd31f7b3fd403b2f7cb89e1171ffc0aad9d293f1715448daf4d

Observation 2b3a3ea5-6b82-430b-a2e1-714f927cb42b · inbound

Qwen-Music Technical Report cites this paper.

Qwen-Music Technical Report Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T03:47:22.936776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T03:47:22.936776Z digest=sha256:aef6662c9f78c07476a0899d721b549033abc6e76cc6df21e2a52657a40606c5

Observation 00a124f4-2a01-4bab-8e06-eb59a87321a6 · inbound

Qwen-Music Technical Report cites this paper.

Qwen-Music Technical Report Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T06:52:27.502401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:52:27.502401Z digest=sha256:105c2cc98f6a3f784c39c3019f190b50282c1bdd78b02df0a8a4f4eb20aa484c

Observation ede239d2-b610-4e20-a7c0-d065cdc4d680 · inbound

FlowSonic: Stable Zero-Shot Music Editing via High-Order Trajectory Integration cites this paper.

FlowSonic: Stable Zero-Shot Music Editing via High-Order Trajectory Integration Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T17:47:12.237390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:47:12.237390Z digest=sha256:3ea55fef6833d238340fcbd539fc2c853e1951d3a94ec9e2714e2fa0169c77f5

Observation ab0c43f7-3f42-4f3f-8ea9-1629982ca015 · inbound

MusiChat: Vibe Composing for Music Creation cites this paper.

MusiChat: Vibe Composing for Music Creation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-31T23:25:26.807855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:25:26.807855Z digest=sha256:3d16363a2c849159438d97ab48f28c4e7953a70d7fefa1bd28f07593a6785c92

Observation b2ddd9e8-d2de-47c3-ad66-9185a88d98ef · inbound

Where Does AI Innovation Go? Measuring Research Attention Imbalance in AI Music cites this paper.

Where Does AI Innovation Go? Measuring Research Attention Imbalance in AI Music Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T18:44:12.731924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:44:12.731924Z digest=sha256:50d3bd3e342cc3a274e0ade5114a1785039d7f06263780b6dbe2fd3fdfccafcf

Observation 628d96e7-bdf5-45d1-89ed-20a6d4887ea2 · inbound

Beyond Reconstruction: Full-Context Generative DiT for Music Generation cites this paper.

Beyond Reconstruction: Full-Context Generative DiT for Music Generation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T04:30:20.070783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:30:20.070783Z digest=sha256:74cab5016554d02cff18018dc829b4fd55e9d1ded2515c3bab5974ba75301d9a

Observation 42d740f0-3ddf-404a-8089-d61eea4afbea · inbound

MusicLayout: Explicit Structural Planning for Controllable Text-to-Music Generation cites this paper.

MusicLayout: Explicit Structural Planning for Controllable Text-to-Music Generation Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T04:19:59.886390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:19:59.886390Z digest=sha256:dcf6aeebb67c4c2f6f9ab87342b20421ed32d19c74b03ff922c8f81473acb343