Pith. sign in

Paper Citation Record · LEDGER

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation

As of 13 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2507.05894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.05894 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:20:44.570826Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ba603bf1-c396-4036-bf79-3f47981a2c36 · outbound

This paper cites Simple and Controllable Music Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Simple and Controllable Music Generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.648942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.648942Z digest=sha256:e786ac63a804efeadda35f3c9ae4cd7c86535d73820970f1a161494153511c91

Observation d4b057d7-75a5-4b56-8fea-0c96954640a9 · outbound

This paper cites an unresolved cited work.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Unresolved cited work

Reference 2

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T19:20:45.153418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-06T19:20:43.740072Z digest=sha256:665d9e1b8180a73ba89e9639f7c93dd685294b333bb6fd7f5a78cf0a2c0a5476

Observation f9add67a-2bfd-4ebf-9749-060b1360764e · outbound

This paper cites LP-MusicCaps: LLM-Based Pseudo Music Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation LP-MusicCaps: LLM-Based Pseudo Music Captioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.802455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.802455Z digest=sha256:31265c2c3df1b486492a46105ccb7183304306274b312a0414f5be6c272c089c

Observation 12320d45-088a-444a-845e-6be3175a7b5b · outbound

This paper cites Gemmeke, Daniel P.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Gemmeke, Daniel P

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:43.904648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:43.904648Z digest=sha256:abcf63618b1327cd1cec67b71afe0ae3f096641da4252b74fdfdef5c45ba8122

Observation 054a047f-9b53-4546-a053-249647af5b9e · outbound

This paper cites an unresolved cited work.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:20:45.334658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-06T19:20:43.993872Z digest=sha256:26c18a873b7caa11d4f7b7063850eff057a675b2b2c804f5faf930e2cfe8c54d

Observation 759400cf-083b-4ff0-92e8-d68d6666c208 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.089294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.089294Z digest=sha256:8704fc49d614f59df1af75599c59245fbdc5e20d6c19468615a0a5d20437f2b8

Observation 07307dde-0908-42db-9131-dcd5f8b00567 · outbound

This paper cites Mixtral of Experts.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Mixtral of Experts

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.174681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.174681Z digest=sha256:cbb97b8ba26a11c5218a821dbb1124f8931e612100bed3b673a65891ac18c6d8

Observation e3a64ba9-8a71-46c1-a870-394386fe63e5 · outbound

This paper cites Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.249296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.249296Z digest=sha256:9574b2eefca148612678ea2cc2843f9239bfcabe3da11377db7fe0ada25a9cac

Observation cb3f769f-f12c-4815-b75c-438ab21490ca · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation AudioGen: Textually Guided Audio Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.308451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.308451Z digest=sha256:fa6b0cf47c312d9419acfc78ce196409d4c3215c52d0f9e4f7ab5d4ca2ade873

Observation 6a77f882-56df-428e-a1ab-501cd04a7761 · outbound

This paper cites MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.402435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.402435Z digest=sha256:b58ac1c54c729e887270ffe46f34455a664b76a4b6c2a8962073c93fb807dc62

Observation f6b4739f-16a9-4b7c-8066-b9f4b34dc5a6 · outbound

This paper cites SwinBERT: End-to-End Transformers with Sparse Attention for Video Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation SwinBERT: End-to-End Transformers with Sparse Attention for Video Captioning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.449219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.449219Z digest=sha256:30e91c3682a0b676ce6259b33df46f65e24a9479951744d2f48581b4193adab5

Observation 10078921-d3a7-4300-a691-7e54e76bed42 · outbound

This paper cites Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.506954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.506954Z digest=sha256:1dd9a6fbaad586ee9a6dcddf117377b9f224e486945be77de56884cec9a63312

Observation eb1146fc-1f1a-4acb-aaed-be04486b0cf8 · outbound

This paper cites Diffsound: Discrete Diffusion Model for Text-to-sound Generation.

MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation Diffsound: Discrete Diffusion Model for Text-to-sound Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:44.570826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:44.570826Z digest=sha256:2a190d91d0d7a4d8a9678e77a2f6e9865aea1705eaf6a9b5f43c21644688337c

Pith citing papers

No inbound Pith citation observations are available.