Pith. sign in

Paper Citation Record · LEDGER

LP-MusicCaps: LLM-Based Pseudo Music Captioning

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2307.16372.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.16372 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:16:20.541373Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T19:25:32.407848Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1c11e66d-2ce8-47fa-8e78-4dbe6177f765 · inbound

A Survey of Hallucination in Large Foundation Models cites this paper.

A Survey of Hallucination in Large Foundation Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:21:00.891169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-16T15:21:00.778049Z digest=sha256:b400d5183f3e1123e23156c2bc1eafc23c5cf86a29022dc3f9d713fbb9f5ed8c

Observation 71e830fe-1dcb-470f-a412-9c130f7eb41e · inbound

SALMONN: Towards Generic Hearing Abilities for Large Language Models cites this paper.

SALMONN: Towards Generic Hearing Abilities for Large Language Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:29:46.379213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-18T02:29:46.242983Z digest=sha256:8b57a7929fef85ce58c2e127355e0cfa30efa834413895693ee4cbf241a85ccc

Observation 89a90ebf-0e39-4f7c-9c86-553f337d8a4d · inbound

Do Captioning Metrics Reflect Music Semantic Alignment? cites this paper.

Do Captioning Metrics Reflect Music Semantic Alignment? LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T18:16:20.541373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:16:20.541373Z digest=sha256:7091574fda9b897f10daf22932ec7c0467a943b3531e21583097b080faecb4b8

Observation e601f507-1b25-4895-bf77-6822b18f1367 · inbound

Improving Controllability and Editability for Pretrained Text-to-Music Generation Models cites this paper.

Improving Controllability and Editability for Pretrained Text-to-Music Generation Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:22:17.865512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:22:17.865512Z digest=sha256:39e8abe61635db4aee49f42d0af00e973888d65244e597b63f2d438f62db7c17

Observation ef3c6ed6-235b-405c-8a20-632aec7bb33a · inbound

HunyuanVideo: A Systematic Framework For Large Video Generative Models cites this paper.

HunyuanVideo: A Systematic Framework For Large Video Generative Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T07:42:43.525030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-23T07:41:58.617477Z digest=sha256:69c3027c1b386979cda61ce4f2b6d5e9fcfaac1882872605aa9652d7aab6c66d

Observation ef2a7171-98f1-4fb8-842f-c7842861de31 · inbound

Combining Genre Classification and Harmonic-Percussive Features with Diffusion Models for Music-Video Generation cites this paper.

Combining Genre Classification and Harmonic-Percussive Features with Diffusion Models for Music-Video Generation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T20:29:54.373108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:29:54.373108Z digest=sha256:50f96bad8449d737e83e51b34d53f08a89018c11290f2424914a1638ccb56348

Observation 3426ff5a-082f-44e3-a0a9-3a0b7ff33c0e · inbound

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey cites this paper.

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T14:59:01.537265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:59:01.537265Z digest=sha256:c2dd24a1236f4045b91cf4c84c1678a452b0ca67df030a8da496dceed61e1122

Observation 50b30830-190f-4d5d-8ae1-efca20564ce1 · inbound

ETTA: Elucidating the Design Space of Text-to-Audio Models cites this paper.

ETTA: Elucidating the Design Space of Text-to-Audio Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T00:45:19.924941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:45:19.924941Z digest=sha256:3de313688767281424ce48fc0b077915d001ce8570dc11745ecfd5cf0c991532

Observation ac1e5441-76a6-4da7-a100-9884517d56c7 · inbound

One Does Not Simply Meme Alone: Evaluating Co-Creativity Between LLMs and Humans in the Generation of Humor cites this paper.

One Does Not Simply Meme Alone: Evaluating Co-Creativity Between LLMs and Humans in the Generation of Humor LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T18:19:03.849477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:19:03.849477Z digest=sha256:8444d1adbcc4ebc23bf0097baf6435fe2b8f788668d85cae5ce3c21410cbd248

Observation a1082ab1-6e3e-469e-8875-69440584b069 · inbound

Qwen2.5-Omni Technical Report cites this paper.

Qwen2.5-Omni Technical Report LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T17:54:03.298059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T17:54:03.225439Z digest=sha256:592fcd869cd16a687335de9b18e9fab91d37b4203599ffba4cdfb3dcd7d89f59

Observation 52ae04ab-437f-4c74-a48b-f8a1ae285d26 · inbound

NoRe: Augmenting Journaling Experience with Generative AI for Music Creation cites this paper.

NoRe: Augmenting Journaling Experience with Generative AI for Music Creation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:50:36.461153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:50:36.461153Z digest=sha256:2c0b5ed2eabbd68f1e44a8d7d9e91fc5ce535dcc43c5e65017472cde76bfda52

Observation 18e95974-ec41-4849-a687-64572867160f · inbound

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections cites this paper.

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T00:49:47.415707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:49:47.415707Z digest=sha256:05af6f36c4028c37098f856e46d96f1c665de19f546d655a3c82d0c8f1781244

Observation 44a70889-5bed-41d9-b16c-75ec9af2d7b7 · inbound

Do Music Preferences Reflect Cultural Values? A Cross-National Analysis Using Music Embedding and World Values Survey cites this paper.

Do Music Preferences Reflect Cultural Values? A Cross-National Analysis Using Music Embedding and World Values Survey LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:05.789790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:05.789790Z digest=sha256:06df4bd7e52920cdaf37a21f76b8653fed60e7f576ba6560e31bc4bfd2269cbb

Observation 2d340b80-8165-4d77-9e8f-7a0c1afe2953 · inbound

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment cites this paper.

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:20:49.195794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-22T00:18:57.564840Z digest=sha256:f0ce7a5d1b0042b742ab7f2454dfe4ec57f3f4f233a2e37c88ba8e7a0183aad9

Observation f9add67a-2bfd-4ebf-9749-060b1360764e · inbound

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation cites this paper.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.802455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.802455Z digest=sha256:1f0a2a0156b825839f3606d9cbbec4686837b2f984d2596ae713b53b02fcf698

Observation aaa14029-16e7-4e00-8971-645909f6fafd · inbound

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models cites this paper.

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:42:44.904068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T03:42:44.523919Z digest=sha256:8b05e0b84e65130be338e14223f37858f386259e156494beff6eda60b6f4c3cd

Observation a5be9aea-0026-43be-80c2-0534256b5b45 · inbound

Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos cites this paper.

Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:02:46.431246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:02:46.431246Z digest=sha256:89e63cfeceecc4c85dd0b6554f595b34915867efe3060173a5da21f91bc0db4b

Observation dc40ba76-b226-4257-adbc-9681683df35e · inbound

TinyMU: A Compact Audio-Language Model for Music Understanding cites this paper.

TinyMU: A Compact Audio-Language Model for Music Understanding LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:17:36.946004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T08:17:23.740979Z digest=sha256:d21710ba4d9b5685e5524ebd45d00e2714041d6def7277ec5ee26135c9f558c7

Observation bcd8d879-6714-4f64-813f-82606da3d79c · inbound

Reddit2Deezer: A Scalable Dataset for Real-World Grounded Conversational Music Recommendation cites this paper.

Reddit2Deezer: A Scalable Dataset for Real-World Grounded Conversational Music Recommendation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:31:19.468474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-12T03:30:53.788986Z digest=sha256:bae4788be9baebf498e057299f24b28354c3f945af02860b83759376baba3369

Observation 6197b704-0062-4fec-a798-aa22af34f715 · inbound

Text2Score: Generating Sheet Music From Textual Prompts cites this paper.

Text2Score: Generating Sheet Music From Textual Prompts LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:42:53.775646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T19:40:38.232446Z digest=sha256:640d7fb39a5a039473dc6b72b9396be7cd0c2642400501abaf532c5602d2a1e5

Observation 9b744f3a-ef50-4d36-8984-ea587fa86ac4 · inbound

Text2Score: Generating Sheet Music From Textual Prompts cites this paper.

Text2Score: Generating Sheet Music From Textual Prompts LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T14:10:39.762416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:10:39.762416Z digest=sha256:89820c3ef5fcdb24df187a055850968d633d1da2cadd040ee2169e163271ce6f

Observation 9984ac53-74ab-4ec0-8611-b25820e6926d · inbound

Toward Native Multimodal Modeling: A Roadmap cites this paper.

Toward Native Multimodal Modeling: A Roadmap LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 167

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:04:02.079804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T22:58:38.610609Z digest=sha256:b3a507b5a4dd7e096412f8bfa5f39a12f16af31d131225798fc19d1514734928

Observation cdc22be6-77c5-4a31-b829-d0f92c0dca2a · inbound

Dasheng AudioGen: A Unified Model for Generating Coherent Audio Scenes from Text cites this paper.

Dasheng AudioGen: A Unified Model for Generating Coherent Audio Scenes from Text LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T11:03:18.755161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T10:56:00.765400Z digest=sha256:589cf6ff2a0a665856afaa59dd700a2ca7acb1e72eb7dfd6b3b33f104e47e11f

Observation 54d738a6-5f29-4837-b50b-d5b9e84cea8d · inbound

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models cites this paper.

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T11:55:43.023632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-01T03:31:42.611472Z digest=sha256:017623fa6006a913bad253a18f67bf10f9a8d7edcc50d33dd0812663af5c29fe

Observation 3773e809-0222-4f0c-a62d-792613c90493 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.475069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:3f9f0f05783ae54313ce5273174858aab0a921196a4caa30162b874aded75477

Observation 55803767-6978-4b59-a606-c2e83bd1a72b · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:9cc2c70218f2a5f5944428320bf3d86587b7085a55a3238c8c7211a0a04ec93d

Observation 62e5769b-1db1-4242-b4a5-304f1d3eec2a · inbound

Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking cites this paper.

Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-08T19:25:32.409284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-08T19:18:56.334254Z digest=sha256:938e2ff657ee0af40003507daf3e351f0b7958b2e1bfd28ec8081ccb785d512a