Pith. sign in

Paper Citation Record · LEDGER

LP-MusicCaps: LLM-Based Pseudo Music Captioning

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2307.16372.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.16372 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:16:20.541373Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T19:25:32.407848Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1c11e66d-2ce8-47fa-8e78-4dbe6177f765 · inbound

A Survey of Hallucination in Large Foundation Models cites this paper.

A Survey of Hallucination in Large Foundation Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:21:00.891169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-16T15:21:00.778049Z digest=sha256:b293a01729178b81ad86f3651b5e0f4ade8468df520ea189a18ed068091a604f

Observation 71e830fe-1dcb-470f-a412-9c130f7eb41e · inbound

SALMONN: Towards Generic Hearing Abilities for Large Language Models cites this paper.

SALMONN: Towards Generic Hearing Abilities for Large Language Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:29:46.379213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-18T02:29:46.242983Z digest=sha256:ec32e39100f077d2998f393ff691d31548ae338a276540ff8923e9af87a8da52

Observation 89a90ebf-0e39-4f7c-9c86-553f337d8a4d · inbound

Do Captioning Metrics Reflect Music Semantic Alignment? cites this paper.

Do Captioning Metrics Reflect Music Semantic Alignment? LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T18:16:20.541373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:16:20.541373Z digest=sha256:b5163bcd381695cb73a22a942118cb8054693a183630115cdc90afe039c907bc

Observation e601f507-1b25-4895-bf77-6822b18f1367 · inbound

Improving Controllability and Editability for Pretrained Text-to-Music Generation Models cites this paper.

Improving Controllability and Editability for Pretrained Text-to-Music Generation Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:22:17.865512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:22:17.865512Z digest=sha256:7e31108123127fa42e1e331846356608a113c4d810a5c577f8bcb21525e0f9a0

Observation ef3c6ed6-235b-405c-8a20-632aec7bb33a · inbound

HunyuanVideo: A Systematic Framework For Large Video Generative Models cites this paper.

HunyuanVideo: A Systematic Framework For Large Video Generative Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T07:42:43.525030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-23T07:41:58.617477Z digest=sha256:728f63fea5993ede4ed2278de127b7faea8768c1506d7be3e94e2f8fcd96f394

Observation ef2a7171-98f1-4fb8-842f-c7842861de31 · inbound

Combining Genre Classification and Harmonic-Percussive Features with Diffusion Models for Music-Video Generation cites this paper.

Combining Genre Classification and Harmonic-Percussive Features with Diffusion Models for Music-Video Generation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T20:29:54.373108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:29:54.373108Z digest=sha256:5c0693e77f62b00d252cbae83a574067977ee800c234c223081c4964559f9515

Observation 3426ff5a-082f-44e3-a0a9-3a0b7ff33c0e · inbound

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey cites this paper.

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T14:59:01.537265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:59:01.537265Z digest=sha256:23f2007f7777148fe5848a55adc0ba46d07c94a5999f39349c65626e1bed81e9

Observation 50b30830-190f-4d5d-8ae1-efca20564ce1 · inbound

ETTA: Elucidating the Design Space of Text-to-Audio Models cites this paper.

ETTA: Elucidating the Design Space of Text-to-Audio Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T00:45:19.924941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:45:19.924941Z digest=sha256:60e5a874909f24ab7f45704d664de57f7490f10638ca404186d489d05cba01f5

Observation ac1e5441-76a6-4da7-a100-9884517d56c7 · inbound

One Does Not Simply Meme Alone: Evaluating Co-Creativity Between LLMs and Humans in the Generation of Humor cites this paper.

One Does Not Simply Meme Alone: Evaluating Co-Creativity Between LLMs and Humans in the Generation of Humor LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T18:19:03.849477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:19:03.849477Z digest=sha256:3188a6ed5eece2d24c74174ce50d01dbae358fc225f11e02e2b72f60d5e56291

Observation a1082ab1-6e3e-469e-8875-69440584b069 · inbound

Qwen2.5-Omni Technical Report cites this paper.

Qwen2.5-Omni Technical Report LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T17:54:03.298059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T17:54:03.225439Z digest=sha256:d46821c971a7bd9a8ac790685332197e9f87b22bcfc4ebe91d7884acb79ccfc7

Observation 52ae04ab-437f-4c74-a48b-f8a1ae285d26 · inbound

NoRe: Augmenting Journaling Experience with Generative AI for Music Creation cites this paper.

NoRe: Augmenting Journaling Experience with Generative AI for Music Creation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:50:36.461153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:50:36.461153Z digest=sha256:d47a99b9b5ce29919e86ef4f9d29a78d5e4ba302ee58fb3093e930797ca08b2c

Observation 18e95974-ec41-4849-a687-64572867160f · inbound

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections cites this paper.

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T00:49:47.415707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:49:47.415707Z digest=sha256:2fe6b34d929962f9cf38bf8744f7f4b7309fad6ccbaeffadf8be5b54849bbbd9

Observation 44a70889-5bed-41d9-b16c-75ec9af2d7b7 · inbound

Do Music Preferences Reflect Cultural Values? A Cross-National Analysis Using Music Embedding and World Values Survey cites this paper.

Do Music Preferences Reflect Cultural Values? A Cross-National Analysis Using Music Embedding and World Values Survey LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:05.789790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:05.789790Z digest=sha256:21b0b33d98c26cc0507a9d9b03b37778e8cf81ee97f9cd8f45dfd5b1ee43034e

Observation 2d340b80-8165-4d77-9e8f-7a0c1afe2953 · inbound

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment cites this paper.

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:20:49.195794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-22T00:18:57.564840Z digest=sha256:f366b9883e3faf32c6e7b282724ea835eb41bd1c6375fc5c8629458ef97d0bc3

Observation f9add67a-2bfd-4ebf-9749-060b1360764e · inbound

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation cites this paper.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.802455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.802455Z digest=sha256:15cf7f3867fd3dd209bd9a5d97caedac3c3c6d309b89bd9a4b45b0eb84cefc02

Observation aaa14029-16e7-4e00-8971-645909f6fafd · inbound

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models cites this paper.

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:42:44.904068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T03:42:44.523919Z digest=sha256:df1a3f04f55a55f9d58736993d9dd31760b51b4a70990a7b0eb8f7b794de5b25

Observation a5be9aea-0026-43be-80c2-0534256b5b45 · inbound

Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos cites this paper.

Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:02:46.431246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:02:46.431246Z digest=sha256:fe31cea89a9bd7dcf3315745478c9c5fab541847a0d679229df4ed67b58e8fe7

Observation dc40ba76-b226-4257-adbc-9681683df35e · inbound

TinyMU: A Compact Audio-Language Model for Music Understanding cites this paper.

TinyMU: A Compact Audio-Language Model for Music Understanding LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:17:36.946004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T08:17:23.740979Z digest=sha256:7481c78170bd605e7d6e638650de67fc64f65fa693bda8015cdf20c1ccb77f52

Observation bcd8d879-6714-4f64-813f-82606da3d79c · inbound

Reddit2Deezer: A Scalable Dataset for Real-World Grounded Conversational Music Recommendation cites this paper.

Reddit2Deezer: A Scalable Dataset for Real-World Grounded Conversational Music Recommendation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:31:19.468474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-12T03:30:53.788986Z digest=sha256:9a1d6a88b5964f974545972b705bd58b7e650cd2a789959ae17cdeb7f3783cb0

Observation 6197b704-0062-4fec-a798-aa22af34f715 · inbound

Text2Score: Generating Sheet Music From Textual Prompts cites this paper.

Text2Score: Generating Sheet Music From Textual Prompts LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:42:53.775646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T19:40:38.232446Z digest=sha256:c858642aebff127c001383673d250bdd6e158c9b09cb403afba278885fcc8993

Observation 9b744f3a-ef50-4d36-8984-ea587fa86ac4 · inbound

Text2Score: Generating Sheet Music From Textual Prompts cites this paper.

Text2Score: Generating Sheet Music From Textual Prompts LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T14:10:39.762416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:10:39.762416Z digest=sha256:1e2830d3ff9dafdd78d62a1dc0000b2dcdc868f39e7aff93fa0c7f12497abe5d

Observation 9984ac53-74ab-4ec0-8611-b25820e6926d · inbound

Toward Native Multimodal Modeling: A Roadmap cites this paper.

Toward Native Multimodal Modeling: A Roadmap LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 167

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:04:02.079804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T22:58:38.610609Z digest=sha256:db5aa1792c313581ea22d0d3110a3f3dfa70dd7b65b381623b63ada0c263cfd6

Observation cdc22be6-77c5-4a31-b829-d0f92c0dca2a · inbound

Dasheng AudioGen: A Unified Model for Generating Coherent Audio Scenes from Text cites this paper.

Dasheng AudioGen: A Unified Model for Generating Coherent Audio Scenes from Text LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T11:03:18.755161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T10:56:00.765400Z digest=sha256:251da17ff2870c1b85cbef5774bc2a8c8c1f1deb3a35053e2fef7bc7cc488162

Observation 54d738a6-5f29-4837-b50b-d5b9e84cea8d · inbound

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models cites this paper.

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T11:55:43.023632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-01T03:31:42.611472Z digest=sha256:ce360c50c80a37f088e28220e3276c7ebf1ac34aa27d3438a2db10fa655e900b

Observation 3773e809-0222-4f0c-a62d-792613c90493 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.475069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:87076e354830f15fe6d49201aae1150c1ba01204c6b8eaf5459ff24cea77877f

Observation 55803767-6978-4b59-a606-c2e83bd1a72b · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:ff45944d10d6c924ead26b23d9f98a2f807e37dd2442957f9833967cf1588039

Observation 62e5769b-1db1-4242-b4a5-304f1d3eec2a · inbound

Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking cites this paper.

Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-08T19:25:32.409284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-08T19:18:56.334254Z digest=sha256:dae804fbb036476c6a92f301063acfc0575cbdc5bcf187fc8fdeb15f2bbc85a1