Pith. sign in

Paper Citation Record · LEDGER

MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.12957.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.12957 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:21:55.658153Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T13:28:18.797278Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f64a50f9-8283-4b54-a678-867818026ec2 · inbound

AudioGenie: A Training-Free Multi-Agent Framework for Diverse Multimodality-to-Multiaudio Generation cites this paper.

AudioGenie: A Training-Free Multi-Agent Framework for Diverse Multimodality-to-Multiaudio Generation MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:55.658153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:21:55.658153Z digest=sha256:3cd2c21ddef220ede7bf675681394c22aef592fd029a7897495aeaf233dfddce

Observation 92882141-7d2d-423f-9e38-e178bfac7bf1 · inbound

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections cites this paper.

Video-Guided Text-to-Music Generation Using Public Domain Movie Collections MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:49:43.716277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:49:43.716277Z digest=sha256:729947d7c4015aba1b29b0f9b48d990db8c1e4328504004978e7f3cd6e7c509b

Observation 570ab3ab-09c6-4f44-a80b-eb1002082a44 · inbound

EXPOTION: Facial Expression and Motion Control for Multimodal Music Generation cites this paper.

EXPOTION: Facial Expression and Motion Control for Multimodal Music Generation MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T19:40:03.310979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:40:03.310979Z digest=sha256:57e2cffce7cf894e9fdbe29b8366f5567e85bfe8163f9b0b1c0e5cda2d397b40

Observation b1c89c59-1a2a-4c02-bf83-afc3c1a059e2 · inbound

Controllable Video-to-Music Generation with Multiple Time-Varying Conditions cites this paper.

Controllable Video-to-Music Generation with Multiple Time-Varying Conditions MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T13:31:11.375359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:31:11.375359Z digest=sha256:525e2fd3e3d79c944d0c8103f3d6043ab41d311a6b2aa47a0cdaa6fa6c03bcdd

Observation 00bede47-b91c-4548-92a9-1b49e45abfc6 · inbound

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions cites this paper.

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T00:56:24.684395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T13:10:29.213917Z digest=sha256:32d5ee47dcab94670d1786b235cc08cd03c3d313b35106eb56d328c454895f7a

Observation afaac524-bfe9-4f50-8cb7-1b4c376d4569 · inbound

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation cites this paper.

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:28:18.798650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T08:04:48.283908Z digest=sha256:c629101db2a48e4da90e63eb32096b1780947f9fb690ff42dd58549e984c7165