Pith. sign in

Paper Citation Record · LEDGER

Draw an Audio: Leveraging Multi-Instruction for Video-to-Audio Synthesis

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2409.06135.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.06135 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:22:12.009729Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T22:20:42.074913Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 075061cb-cbf6-47c0-b40f-737cffdf4b44 · inbound

Gotta Hear Them All: Towards Sound Source Aware Audio Generation cites this paper.

Gotta Hear Them All: Towards Sound Source Aware Audio Generation Draw an Audio: Leveraging Multi-Instruction for Video-to-Audio Synthesis

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T14:22:12.009729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:22:12.009729Z digest=sha256:1c5c1132fae26c5e63f5b35a20a5f52f5af46588462268303f16763c5715727b

Observation 78196977-d827-4572-acfa-c5b345d24883 · inbound

YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls cites this paper.

YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls Draw an Audio: Leveraging Multi-Instruction for Video-to-Audio Synthesis

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T17:17:49.319212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:17:49.319212Z digest=sha256:c45933d225fd2408b172c09d3e9895882b40343d221929fa94b1e8457c5cf791

Observation e8c34cb8-e5e7-49d9-b3fb-412f0dd77e56 · inbound

AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation cites this paper.

AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation Draw an Audio: Leveraging Multi-Instruction for Video-to-Audio Synthesis

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-11T11:38:08.683531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:38:08.683531Z digest=sha256:a6a2430fdede3e3cc70d32b595b95a998a303a4fe8c69967d45c36db61062004

Observation b700647e-40d4-4a20-a211-7cf0dce80133 · inbound

MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis cites this paper.

MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis Draw an Audio: Leveraging Multi-Instruction for Video-to-Audio Synthesis

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T11:35:57.851941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:35:57.851941Z digest=sha256:b0acca9c14124b8bf0a95707169b3fbe14ab2964236446db670d129bc0a69ec2

Observation 45628ebe-7b54-4095-ac58-b4fb2c0514cc · inbound

Training-Free Multimodal Guidance for Video to Audio Generation cites this paper.

Training-Free Multimodal Guidance for Video to Audio Generation Draw an Audio: Leveraging Multi-Instruction for Video-to-Audio Synthesis

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T22:20:42.076693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-21T22:17:42.348945Z digest=sha256:eb418959054556ae30ede04977f7381a1ec951b3f22b4d73d1491692315d8c35