Pith. sign in

Paper Citation Record · LEDGER

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation

As of 23 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2507.05894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.05894 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:20:44.570826Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ba603bf1-c396-4036-bf79-3f47981a2c36 · outbound

This paper cites Simple and Controllable Music Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Simple and Controllable Music Generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.648942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.648942Z digest=sha256:6bf0d6c61603d1e0547d64b54d66721cadb4bd2074e59429b34670456e594b5c

Observation d4b057d7-75a5-4b56-8fea-0c96954640a9 · outbound

This paper cites an unresolved cited work.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Unresolved cited work

Reference 2

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T19:20:45.153418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T19:20:43.740072Z digest=sha256:6963326757571d23431e7257cac25a5c05220a39225594eb10c463dc7658705e

Observation f9add67a-2bfd-4ebf-9749-060b1360764e · outbound

This paper cites LP-MusicCaps: LLM-Based Pseudo Music Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.802455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.802455Z digest=sha256:123bbe6eec99efb1c93b67b6fa54bb586752b46651c01cde2ee2e8fd476c799c

Observation 12320d45-088a-444a-845e-6be3175a7b5b · outbound

This paper cites Gemmeke, Daniel P.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Gemmeke, Daniel P

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.904648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.904648Z digest=sha256:746d994281322cf42842dca72c68c164bf9d45961de6754023e9dac5f71ca267

Observation 054a047f-9b53-4546-a053-249647af5b9e · outbound

This paper cites an unresolved cited work.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:20:45.334658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T19:20:43.993872Z digest=sha256:1c04d5b87775a9acb26ba35a31a9cc9169669de87ae716cd43ac00a13fc81ced

Observation 759400cf-083b-4ff0-92e8-d68d6666c208 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.089294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.089294Z digest=sha256:a9a5905e27f967e2c3e7a92b4860062148cf591fa8a375eb78c3e958e59142a3

Observation 07307dde-0908-42db-9131-dcd5f8b00567 · outbound

This paper cites Mixtral of Experts.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Mixtral of Experts

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.174681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.174681Z digest=sha256:62e2dc545b84821efc9555f6c0025316724afa1283ade706a21ba2cd0367aa15

Observation e3a64ba9-8a71-46c1-a870-394386fe63e5 · outbound

This paper cites Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.249296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.249296Z digest=sha256:cf7f783cb5c50f9f0184aefc3ded02265ecf4c0da53b2e938800046516cd6c09

Observation cb3f769f-f12c-4815-b75c-438ab21490ca · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation AudioGen: Textually Guided Audio Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.308451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.308451Z digest=sha256:d0cc0672258e9e94bc08ebe90a16e040c3d1e00eb07ee2726f6e12af36640d3e

Observation 6a77f882-56df-428e-a1ab-501cd04a7761 · outbound

This paper cites MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.402435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.402435Z digest=sha256:300837bab70d06f419a697850adf94018c114bb58e44fb25f090a5d709f82112

Observation f6b4739f-16a9-4b7c-8066-b9f4b34dc5a6 · outbound

This paper cites SwinBERT: End-to-End Transformers with Sparse Attention for Video Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation SwinBERT: End-to-End Transformers with Sparse Attention for Video Captioning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.449219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.449219Z digest=sha256:2f3c904c488339f15899ff8cfe605792dd8f1e8151eeee4ec87aaba2c66d4dfb

Observation 10078921-d3a7-4300-a691-7e54e76bed42 · outbound

This paper cites Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.506954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.506954Z digest=sha256:f22dd00aa6220479505740f7a357897296c34208e8dc082c7a60eb611258ad3f

Observation eb1146fc-1f1a-4acb-aaed-be04486b0cf8 · outbound

This paper cites Diffsound: Discrete Diffusion Model for Text-to-sound Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Diffsound: Discrete Diffusion Model for Text-to-sound Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.570826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.570826Z digest=sha256:60c2b1c5dbe905309e680e5610a25d213470fcc1a044002b9ea0f42c8c228750

Pith citing papers

No inbound Pith citation observations are available.