Pith. sign in

Paper Citation Record · LEDGER

Video-to-Audio Generation with Hidden Alignment

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2407.07464.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.07464 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:02:31.121546Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:56:39.412382Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 07b7ac75-c128-441c-a497-5f555665cb84 · inbound

Movie Gen: A Cast of Media Foundation Models cites this paper.

Movie Gen: A Cast of Media Foundation Models Video-to-Audio Generation with Hidden Alignment

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:26.132308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T14:16:18.521699Z digest=sha256:8c1e54ebec0dbe498efeb08439ab8ace00a2d70e9e973108e3e90aa8896286d7

Observation 1c89f49b-6dd4-4868-8384-4f148663444a · inbound

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance cites this paper.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance Video-to-Audio Generation with Hidden Alignment

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.814761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:8ab0b5375e9394e125dff082c4d6ad8696a67d5a43e48314d685ef51f296e201

Observation 74aec35b-5806-40b6-8d49-c687d54f8bac · inbound

AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance cites this paper.

AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance Video-to-Audio Generation with Hidden Alignment

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:20.054612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T06:24:21.611234Z digest=sha256:1dfc2a9796faac7393ad3830da0c5a004128d98cd68a9503e0d55a78c43ac0c4

Observation 3a8a1f20-b2c6-4787-8ae0-41dd02c5fc73 · inbound

Benchmarking Single-Factor Physical Video-to-Audio Generation cites this paper.

Benchmarking Single-Factor Physical Video-to-Audio Generation Video-to-Audio Generation with Hidden Alignment

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.377464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:41:56.917119Z digest=sha256:ab103f7ac392baa593e50bde1c0034ad38a799b16be069f254fcb1d7accd5132

Observation d377455d-877f-4951-8157-bcaeb1446c68 · inbound

Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer cites this paper.

Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer Video-to-Audio Generation with Hidden Alignment

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.296619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T21:17:48.886421Z digest=sha256:686fbf0845983226cb6109d57f43ef7cb1b16ef9253e4f6c18d75b36e769d515

Observation 74be9e54-8746-4131-b556-385d7d6d9475 · inbound

Foley-Omni: A Unified Multimodal Generation Model from Task-Level Audio Synthesis to Complete Video Soundtrack Generation cites this paper.

Foley-Omni: A Unified Multimodal Generation Model from Task-Level Audio Synthesis to Complete Video Soundtrack Generation Video-to-Audio Generation with Hidden Alignment

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T04:56:39.413881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T08:38:44.906160Z digest=sha256:84e1748d3c547a89aad147113963d5fec00dd44b2c81773f80cc0eb816c58063

Observation 66eed211-a3b8-45f9-88d3-94138eb509cb · inbound

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation cites this paper.

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation Video-to-Audio Generation with Hidden Alignment

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T14:02:31.121546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:02:31.121546Z digest=sha256:2a89ce164d6f34bcd886a5c46adc48ae879838d2e951cc35f2c4f947236a4f15