Pith. sign in

Paper Citation Record · LEDGER

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation

As of 10 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2507.05894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.05894 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:20:44.570826Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ba603bf1-c396-4036-bf79-3f47981a2c36 · outbound

This paper cites Simple and Controllable Music Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Simple and Controllable Music Generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.648942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.648942Z digest=sha256:1cd779a1df5b20ac7f229f8658bc4c829c59218203f6d618268d31bcf371d1a8

Observation d4b057d7-75a5-4b56-8fea-0c96954640a9 · outbound

This paper cites an unresolved cited work.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Unresolved cited work

Reference 2

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T19:20:45.153418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T19:20:43.740072Z digest=sha256:8150373ea6a106f4811f0b689305cd7ca9f76432c2864bf04d8fd9968b97e5b5

Observation f9add67a-2bfd-4ebf-9749-060b1360764e · outbound

This paper cites LP-MusicCaps: LLM-Based Pseudo Music Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.802455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.802455Z digest=sha256:e9b8e02f22b30b57d49f289797e2657ecc0d891ba55fa4625b0f08719a63844b

Observation 12320d45-088a-444a-845e-6be3175a7b5b · outbound

This paper cites Gemmeke, Daniel P.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Gemmeke, Daniel P

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.904648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.904648Z digest=sha256:5d92f42240cd87eb85f929e00acff51f7d37105726bcd546760af3bfe3d92f96

Observation 054a047f-9b53-4546-a053-249647af5b9e · outbound

This paper cites an unresolved cited work.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:20:45.334658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T19:20:43.993872Z digest=sha256:5aea0e9304658610d8fd8e5de04de4be5de01d7a09212c833733bffc59a60bd0

Observation 759400cf-083b-4ff0-92e8-d68d6666c208 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.089294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.089294Z digest=sha256:b4b042c52b1f44007812af72bd9b52f64f04a341a51a71b9fddf580588e8d657

Observation 07307dde-0908-42db-9131-dcd5f8b00567 · outbound

This paper cites Mixtral of Experts.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Mixtral of Experts

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.174681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.174681Z digest=sha256:b3f3ddfffe1cb9e3cc705ece522ed781eab994b6f5b4987b65734b799ca48b23

Observation e3a64ba9-8a71-46c1-a870-394386fe63e5 · outbound

This paper cites Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.249296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.249296Z digest=sha256:468df0a1689261d275970aaea344d213ad617ec371814e9822cbe4b5f1ec8469

Observation cb3f769f-f12c-4815-b75c-438ab21490ca · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation AudioGen: Textually Guided Audio Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.308451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.308451Z digest=sha256:5ffc56aed04216f06b722ac4ebd2d9dd4b7403032276dca7eced74a536f7a5e7

Observation 6a77f882-56df-428e-a1ab-501cd04a7761 · outbound

This paper cites MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.402435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.402435Z digest=sha256:6e2c32f57ce9b9301dbbbdb44dda66c8d716d91a91514978f3d61962d48d8b87

Observation f6b4739f-16a9-4b7c-8066-b9f4b34dc5a6 · outbound

This paper cites SwinBERT: End-to-End Transformers with Sparse Attention for Video Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation SwinBERT: End-to-End Transformers with Sparse Attention for Video Captioning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.449219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.449219Z digest=sha256:b43abf87713d847d319c3b08f2686736b8d60ce535a286ac55f78ba473ad2e94

Observation 10078921-d3a7-4300-a691-7e54e76bed42 · outbound

This paper cites Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.506954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.506954Z digest=sha256:b0ef3be117d06f489d2e8fd959ace049ca721e20360270e2aaeca785ef9e2a93

Observation eb1146fc-1f1a-4acb-aaed-be04486b0cf8 · outbound

This paper cites Diffsound: Discrete Diffusion Model for Text-to-sound Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Diffsound: Discrete Diffusion Model for Text-to-sound Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.570826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.570826Z digest=sha256:628f12f9f4bc6b1c7534eabb9fd0f1f359f4fff47276218bcc2294c4580a794e

Pith citing papers

No inbound Pith citation observations are available.