Pith. sign in

Paper Citation Record · LEDGER

Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2308.11276.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.11276 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:22:08.799825Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T02:29:46.291983Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 412b71da-8b8e-4306-9ab6-499ba218be67 · inbound

SALMONN: Towards Generic Hearing Abilities for Large Language Models cites this paper.

SALMONN: Towards Generic Hearing Abilities for Large Language Models Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:29:46.295934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T02:29:46.242983Z digest=sha256:3aad07d8fa797ef71dfb00ca8970813cfd214f42bfda8e853bbcedf09bfd07a5

Observation 4edca8fd-f310-411c-aeb6-b66e85c338ce · inbound

MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models cites this paper.

MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T19:31:43.897874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:31:43.897874Z digest=sha256:9608c5402bacc5bad77856cda6114348c44dc4709497d2b8e29c241860738a60

Observation 92a23603-9a4f-4c8e-a783-f68636d82b58 · inbound

Exploring GPT's Ability as a Judge in Music Understanding cites this paper.

Exploring GPT's Ability as a Judge in Music Understanding Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T16:23:18.478998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:23:18.478998Z digest=sha256:6a7be62b5a9b8273f44e8b4d42047b8b37c6a4a03396d1c70cc0cdc5f3d2b0b3

Observation b5fad33f-a609-4975-a59e-f0e4cf0cca40 · inbound

Enhancing Non-Core Language Instruction-Following in Speech LLMs via Semi-Implicit Cross-Lingual CoT Reasoning cites this paper.

Enhancing Non-Core Language Instruction-Following in Speech LLMs via Semi-Implicit Cross-Lingual CoT Reasoning Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T05:22:08.799825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:22:08.799825Z digest=sha256:2aab30527c8bdb303e5b5b945d3f706d32b1f7aadf1430542cb3aa5277a91512

Observation 4ce729e8-0eca-4eeb-ae67-420324ba854c · inbound

Selective Invocation for Multilingual ASR: A Cost-effective Approach Adapting to Speech Recognition Difficulty cites this paper.

Selective Invocation for Multilingual ASR: A Cost-effective Approach Adapting to Speech Recognition Difficulty Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:50.833666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:10:50.833666Z digest=sha256:0d98f26ad42bdc74408189f69aed59afa188231f66c972410537b2814810173d

Observation 10078921-d3a7-4300-a691-7e54e76bed42 · inbound

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation cites this paper.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.506954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.506954Z digest=sha256:b28c619de9bcd76da39f9359ae13f3a4b9ec11d04d21bec64f04a97ac32f84db

Observation 52bb2a9c-25da-45f6-b483-6ca5e705639b · inbound

The TEA-ASLP System for Multilingual Conversational Speech Recognition and Speech Diarization in MLC-SLM 2025 Challenge cites this paper.

The TEA-ASLP System for Multilingual Conversational Speech Recognition and Speech Diarization in MLC-SLM 2025 Challenge Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:23.708169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:45:23.708169Z digest=sha256:c0e9cb89aa15baf078275e048b9834760eb8aae249cc5fc6129c1ccd4567dfd3

Observation 191bf5ba-0637-4650-b552-aab216b8652b · inbound

Assessing Factual Music Comprehension in Large Audio Language Models cites this paper.

Assessing Factual Music Comprehension in Large Audio Language Models Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T00:29:46.773378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:29:46.773378Z digest=sha256:3b8be3a6d5ca0ae4b887a862d05d9c6864b9a1654df8860ed86570cae7930617