Pith. sign in

Paper Citation Record · LEDGER

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models

As of 23 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 2 inbound Pith citation observations for arXiv:2506.17886.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17886 v2

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:04:34.985831Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:04:34.724104Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T22:00:21.482883Z

Reference resolution

70 of 70 outbound references displayed

  • verified exact0
  • verified fuzzy55
  • unresolved14
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ed0920a8-c9d3-47c0-a97d-6fc1c8cfe3ae · outbound

This paper cites GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.724104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.724104Z digest=sha256:12dfff0c458a0285e676847083f6371012bba5094f311e8c1f9ae9b975ed939b

Observation cedf3215-e059-48bf-9fd1-408b8bdae032 · outbound

This paper cites an unresolved cited work.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:04:35.793177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.729308Z digest=sha256:f3838d6984babccd95a95f095c4dc85da3e4d2cfd097c8ff5d4e7e920355bf87

Observation 5abc4c99-22ab-4541-9072-8478ddd18aa9 · outbound

This paper cites Using a pretrained latent space op- timized for audio-audio retrieval, we train a generative dif- fusion model conditioned on text to generate audio latent embeddings in this space.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Using a pretrained latent space op- timized for audio-audio retrieval, we train a generative dif- fusion model conditioned on text to generate audio latent embeddings in this space

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.781111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.733508Z digest=sha256:787fc69dd05da302223b1471af5addd309f3d32c417b7b9e3b9b8e7d0f3fe422

Observation 4d67e298-f182-4e44-bbe9-c90d04f5f515 · outbound

This paper cites a rock song.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models a rock song

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T19:04:35.769795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.737920Z digest=sha256:4a443ea1eb234b2607e2e9e4c6fd6c524d211e6f5137176a5455dc0c49560d75

Observation abba59fc-91eb-4551-9ad2-0daff3eacbec · outbound

This paper cites an unresolved cited work.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:04:35.747063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.747172Z digest=sha256:4c6ad74bc33202d7717acd1fbc4ef352c66fd4bddccc258c7ee60943f0efaaae

Observation 6f8aa8cf-5252-435b-8bd6-a0bed5a93190 · outbound

This paper cites does not mean.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models does not mean

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.758579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.742870Z digest=sha256:f94dd94d095d8c09ce9d171fa34b061975224d8f385c2ea2d004f503c94a9f6a

Observation e6af7531-07be-4682-9b72-cab8a855f905 · outbound

This paper cites an unresolved cited work.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.751255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.751255Z digest=sha256:de2817b59dc44c9f056027eec2fc097dcacd37c0e1e4656c0a8fb77615dfd3d9

Observation 783d38eb-a68b-43e7-b466-f877f07707da · outbound

This paper cites Contrastive audio-language learning for music,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Contrastive audio-language learning for music,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.729164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.755496Z digest=sha256:080b4f7a086e13b361a04b9d015357f05e13265ab0eb5eae6c2c7c87d6969175

Observation 3be2fcb2-89a8-4eb7-8f96-117f30c219ed · outbound

This paper cites Mulan: A joint em- bedding of music audio and natural language,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Mulan: A joint em- bedding of music audio and natural language,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.717231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.759497Z digest=sha256:fcdc3afaf2496b5b0ade20e741c2a780611513f0d1fbd945393ba977c0a20682

Observation 1750edf8-a126-4b4f-b69b-f7c43db5410c · outbound

This paper cites Large-scale con- trastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Large-scale con- trastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.705112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.763241Z digest=sha256:aafa28b786b12381d8513eb975432a2dd5ea3989bbb55f83876f00d728f26af8

Observation 131abbe4-db9a-4324-bed0-9510b723969a · outbound

This paper cites Collap: Contrastive long-form language-audio pretraining with musical temporal structure augmentation,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Collap: Contrastive long-form language-audio pretraining with musical temporal structure augmentation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.692681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.767786Z digest=sha256:8b18584c74dd32f06c994d0fd1c3495ea12641928b9089b65e4b905dab40549b

Observation d99bd70a-65c2-4448-8403-760e7fa479d1 · outbound

This paper cites Clap learning audio concepts from natural language supervi- sion,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Clap learning audio concepts from natural language supervi- sion,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.680699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.772208Z digest=sha256:e7178d872b9e773e2a127c82b832824c094c69a5a8d33d097f638dd67dc6179e

Observation b2bd12bd-c2b2-4138-87fb-2b0671e685b7 · outbound

This paper cites Denoising diffusion probabilistic models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Denoising diffusion probabilistic models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.776734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.776734Z digest=sha256:221daaaa0eae670aa7627f6e3e706286e2b9b565f12c344bba9c83d06f3b032d

Observation 3cd4ff71-896a-4e43-af3b-b05b60d65d5c · outbound

This paper cites AudioLDM: Text- to-audio generation with latent diffusion models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models AudioLDM: Text- to-audio generation with latent diffusion models,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.661843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.780501Z digest=sha256:a4cffae7a89664e58248f90be9284cf43c21f1d50fd5035c245624e8d84588a7

Observation 33e189d6-8557-48ac-91b9-4c0d8bfbedb1 · outbound

This paper cites Audioldm 2: Learning holistic audio generation with self-supervised pretrain- ing,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Audioldm 2: Learning holistic audio generation with self-supervised pretrain- ing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.651020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.784453Z digest=sha256:bd4dd4881b94354a19665ba35508d1c7dfce34e8d379d29a236011fe514aa54e

Observation ed749373-da6f-4a8c-9420-d5ebff72d5c0 · outbound

This paper cites MusicLDM: Enhancing Novelty in Text-to-Music Generation Using Beat-Synchronous Mixup Strategies.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models MusicLDM: Enhancing Novelty in Text-to-Music Generation Using Beat-Synchronous Mixup Strategies

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.788182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.788182Z digest=sha256:857be9b05c901b312b56e6d278c5d2afdbf5f00ea9778cb3b973fbaaf7dc1f2b

Observation 39e0218a-2ad5-4e8b-9e36-bee4ec1cd302 · outbound

This paper cites Music con- trolnet: Multiple time-varying controls for music gen- eration,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Music con- trolnet: Multiple time-varying controls for music gen- eration,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.640438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.792073Z digest=sha256:4563b10efd1b252543fb0a410a7a19e2e7703692c7821b86bfa96a7f02eed421

Observation caf6794e-518b-4bad-9b86-efb31fd5377d · outbound

This paper cites Text-to- audio generation using instruction guided latent diffu- sion model,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Text-to- audio generation using instruction guided latent diffu- sion model,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.628563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.795807Z digest=sha256:09d4fa765c1424f5f4e562d0ba01bd911ea73f3ea46d1f52e91783b402e87c2d

Observation c0857d57-fbfd-427a-a6b5-c940aeeef7db · outbound

This paper cites Diff-a-riff: Musical accompaniment co-creation via latent diffu- sion models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Diff-a-riff: Musical accompaniment co-creation via latent diffu- sion models,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.617241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.799474Z digest=sha256:6001ec0361a5bec38402de1a16818f4dd35f7df0005d58a0174f0aa43826c1bd

Observation 72324da7-483d-41ac-828a-c6367c1a5a2f · outbound

This paper cites Ditto: diffusion inference-time t-optimization for mu- sic generation,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Ditto: diffusion inference-time t-optimization for mu- sic generation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.605214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.803544Z digest=sha256:9ca5c6d97d8f1f2f59bf39256a141a2afb26e0dcd3b5d13a9e0cb254e5cf62f9

Observation 81277619-1eed-4019-af56-a064c959082e · outbound

This paper cites Long-form mu- sic generation with latent diffusion,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Long-form mu- sic generation with latent diffusion,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.593145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.807393Z digest=sha256:687dcf948ac2765e93b31de0f4166eb801275027909d88fdd48d88fd22656eef

Observation c0f21343-fca6-4ea1-a2e2-5c0728cbb755 · outbound

This paper cites Fast timing- conditioned latent audio diffusion,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Fast timing- conditioned latent audio diffusion,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.580645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.811170Z digest=sha256:470990fc2d40201d2ab32214c7a39a6a7bcf49abc8afeb2c4b5fe2879c3aa0f3

Observation a8752a35-85d6-4b5b-ada1-e9b2218d7938 · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.814749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.814749Z digest=sha256:381fdaf96cdde1a86c73d6f2776c82b6024eb17d7339d72c5488c9f3b8940672

Observation d492cbcd-5a8d-4ccb-848a-7001b9561939 · outbound

This paper cites Null-text inversion for editing real images using guided diffusion models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Null-text inversion for editing real images using guided diffusion models,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.568602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.818639Z digest=sha256:9fe37dfe6df2c1c236a883a6f88e6aa210fdb60e8b926c9c04118bc3a45c11d6

Observation 90181c40-35a6-485d-b6ff-d4d367efa5ac · outbound

This paper cites Dynamic prompt learning: Addressing cross-attention leakage for text- based image editing,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Dynamic prompt learning: Addressing cross-attention leakage for text- based image editing,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.557569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.822340Z digest=sha256:8f1fa4e6f0e02050a3edd47bad4ce6e0574f6f019ea532af76abbc41246f0733

Observation 7c6b48a7-2097-4783-8212-ea42d6671e06 · outbound

This paper cites Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.825771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.825771Z digest=sha256:f312df7870417a2c0a410cead0899fe034f59ffea912b0972fef2b43c20bfe1a

Observation 3a4e9d24-eeef-41c9-b370-034cc016bf9d · outbound

This paper cites Compositional in- version for stable diffusion models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Compositional in- version for stable diffusion models,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.547180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.829422Z digest=sha256:a580fc6e54485744d84a2dbdc65ebb50d8ab07ebc6417e02f0fd419cea52b2f6

Observation de5fa644-fc81-481c-bbb5-58c64c7c8bb8 · outbound

This paper cites Learn- ing transferable visual models from natural language supervision,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Learn- ing transferable visual models from natural language supervision,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.537188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.833725Z digest=sha256:aa7d17ff2749b12195f1a5c02ae686d2677fbb7a2ac8d41b5b323ca7a37523c0

Observation aadb7bce-42e7-44c3-a84d-fe41170fecdc · outbound

This paper cites Sigmoid loss for language image pre-training,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Sigmoid loss for language image pre-training,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.527054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.837704Z digest=sha256:d61e41785ba10ad1828cbef7cb5ed5088683687ac1f058200fef906739c27f0f

Observation 32ed8994-28c1-459d-84b9-b4e7bb584e19 · outbound

This paper cites Improving fine- grained understanding in image-text pre-training,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Improving fine- grained understanding in image-text pre-training,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.515271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.841097Z digest=sha256:168a200d32f2cea35e576495eb53e5ee073e704378cb3fd46953e8aa9f40292c

Observation d3daaa15-e3e2-43eb-8e5c-f337cc98f713 · outbound

This paper cites A simple framework for contrastive learning of visual represen- tations,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models A simple framework for contrastive learning of visual represen- tations,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.504818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.844695Z digest=sha256:3e8e994121d3777a22e8a031c980c4b3bbb0c0d7eb9b5e73cab64ebc22cb9e33

Observation 1fd7d533-778a-4161-8583-9c8dd7a8d009 · outbound

This paper cites T-clap: Temporal- enhanced contrastive language-audio pretraining,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models T-clap: Temporal- enhanced contrastive language-audio pretraining,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.494856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.848206Z digest=sha256:ff475dcef9b59a2c23f23cba3617b85c77789cdcae96f4ebea42293f0006410f

Observation 7e5b7515-5304-410e-949a-90a8706efc91 · outbound

This paper cites Augment, drop & swap: Improving diversity in llm captions for efficient music-text representation learning,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Augment, drop & swap: Improving diversity in llm captions for efficient music-text representation learning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.483739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.851695Z digest=sha256:d8bce6574b81906185514d10fc781c352158e1c994851d4afe812170cce81933

Observation 342b023c-404f-4bc3-8ab8-f4b657832a22 · outbound

This paper cites Cacophony: An improved contrastive audio-text model,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Cacophony: An improved contrastive audio-text model,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.472403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.855496Z digest=sha256:69bd23cc57c21677d2ec8d819451986d50db893981c8e20955eb8c6aa227b6e7

Observation 7f134a9d-6340-4d63-a1cb-c7980412ad6d · outbound

This paper cites Spec- MaskGIT: Masked Generative Modeling of Audio Spectrograms for Efficient Audio Synthesis and Be- yond,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Spec- MaskGIT: Masked Generative Modeling of Audio Spectrograms for Efficient Audio Synthesis and Be- yond,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.461489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.859067Z digest=sha256:38ccbc84f011c384bff0e26a3069917a8fa32aa8f8fc04785bc641d541bfe427

Observation a4096369-5be7-4155-a002-0fdb216daf30 · outbound

This paper cites Drcap: Decoding clap la- tents with retrieval-augmented generation for zero-shot audio captioning,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Drcap: Decoding clap la- tents with retrieval-augmented generation for zero-shot audio captioning,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.450757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.862709Z digest=sha256:f4ba5aca7de3389a8a6438052d3e2e82627cd4a2986ee1e25f4f1147567696ce

Observation 80880caa-5424-45ab-bd9a-3f07acb7979b · outbound

This paper cites Recap: Retrieval-augmented audio captioning,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Recap: Retrieval-augmented audio captioning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.440153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.866590Z digest=sha256:620c11683472c99b5c1fbe0802f2f6935665855cefda5a42b5011908631920f6

Observation bb1880d1-975d-4e99-a0d1-a7e81a24ed96 · outbound

This paper cites Taming trans- formers for high-resolution image synthesis,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Taming trans- formers for high-resolution image synthesis,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.428626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.870069Z digest=sha256:84f036f9bd08fa3b67cf4fc32d3ed20874624560d1c3e404e8fd7264d16181b2

Observation d3b946b7-b0b0-4de8-97d5-7b797254fe81 · outbound

This paper cites Scaling autoregres- sive models for content-rich text-to-image generation,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Scaling autoregres- sive models for content-rich text-to-image generation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.417117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.873584Z digest=sha256:8e0267a1954340a238d63a3519bd6d9349051c42155fc0f40a9beafcf4f87a9a

Observation e219f080-c386-452f-822a-42ac4141a8d4 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.877039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.877039Z digest=sha256:b52864b6ed59fb15b273468755a371101e851eefca86877c85424289ad622a59

Observation 9e9ace1a-d331-44eb-ac16-0a28a915baa0 · outbound

This paper cites Diffgap: A lightweight diffusion module in contrastive space for bridging cross-model gap,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Diffgap: A lightweight diffusion module in contrastive space for bridging cross-model gap,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.405759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.880838Z digest=sha256:2e906186f65fe383acb434713af9d18ae2ebecac44edd47de900af0d6c7e9234

Observation 4f4e5aae-c4cb-42b8-a6b6-9372b5118a3b · outbound

This paper cites Classifier-free diffusion guid- ance,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Classifier-free diffusion guid- ance,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.394149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.884579Z digest=sha256:04a8b72332187ecde342d480631b2f8a59e304d477132448aa521c19966d543a

Observation 4c13215d-c5b0-477e-93b0-ffc46c719cc7 · outbound

This paper cites Scalable diffusion models with transformers,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Scalable diffusion models with transformers,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.382797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.888240Z digest=sha256:34794e9b0ce2d6b122b8046483ec6129d4adb0e5567c2df59a768101f6851d52

Observation de74626a-dec7-42f2-a063-90920caaaff0 · outbound

This paper cites High- resolution image synthesis with latent diffusion mod- els,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models High- resolution image synthesis with latent diffusion mod- els,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.371444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.891814Z digest=sha256:5c8f51b9f236bdf4e9d0f9c7a597f8236703869a529173ea232f79fa80bc9b51

Observation 13dfd92f-3f0d-49b7-8b81-fcbcaf51db5e · outbound

This paper cites Moûsai: Text-to-music generation with long-context latent dif- fusion,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Moûsai: Text-to-music generation with long-context latent dif- fusion,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.360620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.895451Z digest=sha256:8f5437b8dcc0765efa69ba3e37326d402c3fa604c737c456fef73a6cf86b148d

Observation 6ab6ccb7-8f53-4f51-9dd4-6cbd33db5cfa · outbound

This paper cites Prompt tuning inver- sion for text-driven image editing using diffusion mod- els,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Prompt tuning inver- sion for text-driven image editing using diffusion mod- els,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.349635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.898976Z digest=sha256:dcebe5fa5462e7ab430a7335d5b950fbd0847c7dd75ca383d617b29a45ec553e

Observation 7ebbc67f-4cba-4007-bc86-2bfb35b4a673 · outbound

This paper cites Prompt- to-prompt image editing with cross-attention control,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Prompt- to-prompt image editing with cross-attention control,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.338575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.902627Z digest=sha256:86a441bc7c62b7a5b6c75cdfffbac1966bfdbf876c4052b385445643c7a70f3a

Observation 674dd043-3c3e-4ee9-92af-b16bb6f579fa · outbound

This paper cites Prompt Sliders for Fine-Grained Control, Editing and Erasing of Concepts in Diffusion Models.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Prompt Sliders for Fine-Grained Control, Editing and Erasing of Concepts in Diffusion Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.906186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.906186Z digest=sha256:df9bcff8d307a255897bb2008f9fb99fc820f230e2ee43da6578135898ae0b28

Observation 74d6d795-70bc-458c-9320-8860f6d19f41 · outbound

This paper cites Re- paint: Inpainting using denoising diffusion probabilis- tic models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Re- paint: Inpainting using denoising diffusion probabilis- tic models,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.327988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.910067Z digest=sha256:9842cfcfb44b7729be4a15ff1a7da6b41a8d9cb4265dd13450332d47b3352426

Observation 2b7212e1-ef88-4155-9043-2fd46b586eee · outbound

This paper cites Negative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Negative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.913971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.913971Z digest=sha256:ed1508a238f4b28c2462b232e1e6d2bee337eb68034325e2b05d265393d60342

Observation aaac58ee-a01f-48ab-840d-bdcec4884768 · outbound

This paper cites Musicmagus: zero-shot text-to-music editing via diffusion models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Musicmagus: zero-shot text-to-music editing via diffusion models,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.316632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.917782Z digest=sha256:aa0df68d8d548b136f59d2c779c6aec8f1ec6975aff042cf9803eaccec52fa76

Observation a23a3d75-0cc1-4a3f-8c6c-32480d72e054 · outbound

This paper cites Arrange, inpaint, and refine: steerable long-term music audio generation and editing via content-based controls,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Arrange, inpaint, and refine: steerable long-term music audio generation and editing via content-based controls,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.304947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.921357Z digest=sha256:56bf65e91a9080841efcb22603a608633e912231e45bb9d14ee467810b88b964

Observation c53b2878-8b42-493c-8a84-f74613426e0a · outbound

This paper cites Disentangled multidimensional metric learning for music similarity,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Disentangled multidimensional metric learning for music similarity,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.294205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.925213Z digest=sha256:5e5c2f3c80eb90b465aac818c2e694633fed83d3d5b51b8da91843f4aaabaf24

Observation 91a51648-2e33-4665-b51d-7dd9cb534d52 · outbound

This paper cites Leave-one- equivariant: Alleviating invariance-related information loss in contrastive music representations,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Leave-one- equivariant: Alleviating invariance-related information loss in contrastive music representations,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.284123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.928844Z digest=sha256:c4a7524db407e45d2b9b45dcec8ee5d7adcc0de2c41cea78c919cafd46276489

Observation c3cc4f80-4940-406e-b5db-a3cd8a0e2897 · outbound

This paper cites Similar but faster: manipulation of tempo in music audio em- beddings for tempo prediction and search,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Similar but faster: manipulation of tempo in music audio em- beddings for tempo prediction and search,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.271040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.932222Z digest=sha256:286c77e81a6b11106b6a5a2c39cf833bcd24d61d94d7bf96a5f8d2affb91476f

Observation 4ac1ed93-24f9-452e-b399-270efc1b6833 · outbound

This paper cites Diff4steer: Steer- able diffusion prior for generative music retrieval with semantic guidance,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Diff4steer: Steer- able diffusion prior for generative music retrieval with semantic guidance,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.260053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.935579Z digest=sha256:c05be181c1d72ffc1b5a63031e0297f4c0241393999b308698093eb4175fb0de

Observation 88722465-4711-44b8-bae8-07200c7dae5d · outbound

This paper cites Supervised and unsupervised learning of audio rep- resentations for music understanding,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Supervised and unsupervised learning of audio rep- resentations for music understanding,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.248595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.939280Z digest=sha256:ef46989662132aa70320ce02c82fb7b1629c50f654e9b5b19f41c45d71542194

Observation 3a46cb77-5f7a-4a53-b608-64c95446bd3e · outbound

This paper cites Hts-at: A hierarchical token-semantic audio transformer for sound classifica- tion and detection,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Hts-at: A hierarchical token-semantic audio transformer for sound classifica- tion and detection,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.237503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.942695Z digest=sha256:e786f98d1086d3ce7a7b592d29f9ee7bc5b9680dfe809490e5d3fa30001b3284

Observation e09fd019-bde5-458f-a7c4-63588f14747b · outbound

This paper cites Scal- ing instruction-finetuned language models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Scal- ing instruction-finetuned language models,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.226497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.946057Z digest=sha256:16212d43932a36ee86c06d8493154fae3a6e53806a17bd8b86c65104dd4505fa

Observation f3678114-786e-421f-9ec3-562e73a85f44 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.949482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.949482Z digest=sha256:d29049bd4f1d83a4d3d30ffac251c0cd66774afdab891b5ed82f87e759b228dc

Observation bf07307b-39f4-4a4a-8397-b566ad248545 · outbound

This paper cites Mustango: Toward controllable text-to-music generation,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Mustango: Toward controllable text-to-music generation,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.213436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.953225Z digest=sha256:49ee70cf3cec5fe379266f6cd425d4a0d6a6f80cabd147e5b72e43398d2983cd

Observation e3389da3-07c2-4683-90da-1719eff5af75 · outbound

This paper cites Simple and control- lable music generation,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Simple and control- lable music generation,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.201067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.956709Z digest=sha256:db178827e57e4491722d9e1eddf8cd915b2170a0a9e44f2c2d6a2249e7a2688c

Observation f61cfd5e-2cdf-469e-a0d7-1e7dad83e1d6 · outbound

This paper cites The song describer dataset: a corpus of audio captions for music-and- language evaluation,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models The song describer dataset: a corpus of audio captions for music-and- language evaluation,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.188234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.960138Z digest=sha256:353ab85cbd6ec2462ae39fa41b924a6e56aeeaddbe7c8343b0b0b42ba7d80004

Observation 801c08b4-f68e-482b-9241-0c7439a2cb24 · outbound

This paper cites MusicLM: Generating Music From Text.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models MusicLM: Generating Music From Text

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.963627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.963627Z digest=sha256:cae517e1feb2b9b09389ce461a5935a152a37e8ddfddd613e0376d491f4add88

Observation 7dc546fc-56b0-4e71-bfa1-29ad2eca2864 · outbound

This paper cites Coco-dr: Combating distribution shifts in zero-shot dense retrieval with con- trastive and distributionally robust learning,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Coco-dr: Combating distribution shifts in zero-shot dense retrieval with con- trastive and distributionally robust learning,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.175628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.967500Z digest=sha256:45be32dcbdf4512a890831cad6f91c8250da08dbea80e67ad91538203a34355a

Observation 6751f474-cff9-488c-9565-b66aa8fc1ffb · outbound

This paper cites Selecting which dense retriever to use for zero-shot search,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Selecting which dense retriever to use for zero-shot search,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.164355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.971348Z digest=sha256:d7e8e3a29ee0d9b40cbca8c3f2505d8e1a358b6896ee7fd9f8d6aca3f64bbd86

Observation f1556d10-558e-4fcf-825b-abe4ad3dbff8 · outbound

This paper cites Test-time distribution nor- malization for contrastively learned visual-language models,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Test-time distribution nor- malization for contrastively learned visual-language models,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.152236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.974721Z digest=sha256:5a5687f2191bd821ee16ea7d1905f591a3ff3c08a3db8760cfe9ce0d9ccae6ce

Observation 8bdff5ed-6bff-4522-bad7-fc8d5382834b · outbound

This paper cites Correlation alignment for unsupervised domain adaptation,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models Correlation alignment for unsupervised domain adaptation,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.140764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.978210Z digest=sha256:9d01d1a341f8df352fa32252e5150276f25503deaadba1986846c7ef00f0deda

Observation 4c1bd7a4-d53e-4932-87e2-21397a6aec2c · outbound

This paper cites The Vendi Score: A Diversity Evaluation Metric for Machine Learning.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models The Vendi Score: A Diversity Evaluation Metric for Machine Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.981881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.981881Z digest=sha256:076b69b09c80d2655d200a6fb0301b61e8c3cb2dee20ae3a554e50bcb001802d

Observation 56c9ca0a-23c0-4229-a8a3-cc2b3b3c4897 · outbound

This paper cites The mtg- jamendo dataset for automatic music tagging,.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models The mtg- jamendo dataset for automatic music tagging,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:04:35.128817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T19:04:34.985831Z digest=sha256:8a160aebdc1fe0b13abef9c03099c0b0c9091eba723fda9c9618ae3f3911bbbf

Pith citing papers

Observation ed0920a8-c9d3-47c0-a97d-6fc1c8cfe3ae · inbound

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models cites this paper.

GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T19:04:34.724104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:04:34.724104Z digest=sha256:12dfff0c458a0285e676847083f6371012bba5094f311e8c1f9ae9b975ed939b

Observation 9cf8b871-6fd0-4e2e-8c24-35d7165480c1 · inbound

GaiaFlow: Semantic-Guided Diffusion Tuning for Carbon-Frugal Search cites this paper.

GaiaFlow: Semantic-Guided Diffusion Tuning for Carbon-Frugal Search GD-Retriever: Controllable Generative Text-Music Retrieval with Diffusion Models

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:00:21.488761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T21:58:04.706427Z digest=sha256:221f83a865a999628bb228e461632de19ccbddae07adaad9fc3cd9ce5f5f2aff