Pith. sign in

Paper Citation Record · LEDGER

Video-to-Audio Generation with Hidden Alignment

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2407.07464.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.07464 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:02:31.121546Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:56:39.412382Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 07b7ac75-c128-441c-a497-5f555665cb84 · inbound

Movie Gen: A Cast of Media Foundation Models cites this paper.

Movie Gen: A Cast of Media Foundation Models Video-to-Audio Generation with Hidden Alignment

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:26.132308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T14:16:18.521699Z digest=sha256:2bb3e8b6d8c69826c0a87949cb09d552bbe0734d38557e33b1054d501139f0ce

Observation 1c89f49b-6dd4-4868-8384-4f148663444a · inbound

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance cites this paper.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Video-to-Audio Generation with Hidden Alignment

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.814761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:2dd6a9a21ff22071d053706658d26df1e4370a1e6a9453122f6faa36dc0ca6b7

Observation 74aec35b-5806-40b6-8d49-c687d54f8bac · inbound

AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance cites this paper.

AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance Video-to-Audio Generation with Hidden Alignment

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:20.054612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T06:24:21.611234Z digest=sha256:562eaccada2a1f0fc3e8a4fc4c2c3946b0d77924ac8587e7656d5dde01ba0374

Observation 3a8a1f20-b2c6-4787-8ae0-41dd02c5fc73 · inbound

Benchmarking Single-Factor Physical Video-to-Audio Generation cites this paper.

Benchmarking Single-Factor Physical Video-to-Audio Generation Video-to-Audio Generation with Hidden Alignment

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.377464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T07:41:56.917119Z digest=sha256:3d62a9d464c7a6aaf5eb908a7ce7a9a1bd830b5591c0abbb66b35e24b450536c

Observation d377455d-877f-4951-8157-bcaeb1446c68 · inbound

Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer cites this paper.

Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer Video-to-Audio Generation with Hidden Alignment

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.296619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T21:17:48.886421Z digest=sha256:c048b6220d8ff7ee05e661ce4fcaa057dfae87ee33067fda540fbe5901f3823e

Observation 74be9e54-8746-4131-b556-385d7d6d9475 · inbound

Foley-Omni: A Unified Multimodal Generation Model from Task-Level Audio Synthesis to Complete Video Soundtrack Generation cites this paper.

Foley-Omni: A Unified Multimodal Generation Model from Task-Level Audio Synthesis to Complete Video Soundtrack Generation Video-to-Audio Generation with Hidden Alignment

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T04:56:39.413881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T08:38:44.906160Z digest=sha256:ec9397ca65c0e93775f1304afd1c1374a92421ec1a3d3321962234c82a69d559

Observation 66eed211-a3b8-45f9-88d3-94138eb509cb · inbound

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation cites this paper.

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation Video-to-Audio Generation with Hidden Alignment

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T14:02:31.121546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:02:31.121546Z digest=sha256:2a89ce164d6f34bcd886a5c46adc48ae879838d2e951cc35f2c4f947236a4f15