Pith. sign in

Paper Citation Record · LEDGER

FoleyGen: Visually-Guided Audio Generation

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2309.10537.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2309.10537 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:00:01.471769Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T14:16:25.323927Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ae437815-b30d-4f71-ab37-1ba08bbe4f93 · inbound

Movie Gen: A Cast of Media Foundation Models cites this paper.

Movie Gen: A Cast of Media Foundation Models FoleyGen: Visually-Guided Audio Generation

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:25.365557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-11T14:16:18.521699Z digest=sha256:a4cb7c0de9ec79de11360a32f9cf7075d206d270fa47bd0cba472449237ee628

Observation 7c3068cf-5269-4cb4-977b-ecac4aa4a979 · inbound

Video-Guided Foley Sound Generation with Multimodal Controls cites this paper.

Video-Guided Foley Sound Generation with Multimodal Controls FoleyGen: Visually-Guided Audio Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T12:00:01.471769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:00:01.471769Z digest=sha256:1d97d292513849220081b7e7dbc5b32bcbb73b064b78ab1c91609ec5030b2d3e

Observation a294bb5e-6dbd-4118-9daa-0d89e0471be6 · inbound

SyncFlow: Toward Temporally Aligned Joint Audio-Video Generation from Text cites this paper.

SyncFlow: Toward Temporally Aligned Joint Audio-Video Generation from Text FoleyGen: Visually-Guided Audio Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T23:05:27.408048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:05:27.408048Z digest=sha256:a96f26fbfb0fc231957e7e3d4eacdf5436798a9fc7abc53c6357353dbeafbdc5

Observation 64dbded5-8a80-4a83-85b7-8751e35709b5 · inbound

MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis cites this paper.

MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis FoleyGen: Visually-Guided Audio Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T11:35:57.774813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:35:57.774813Z digest=sha256:af53a4ae397732fc611b43d05b0d702a3e4215648a535651e5370e4de11eba02

Observation a17b413a-662c-4d02-9359-a95abf969829 · inbound

ViSAGe: Video-to-Spatial Audio Generation cites this paper.

ViSAGe: Video-to-Spatial Audio Generation FoleyGen: Visually-Guided Audio Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T01:04:52.823071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:04:52.823071Z digest=sha256:ec45bbc1eca6d66a2435d3e82c35077c40f29ca80eb9a669f92d290ab7a7a066

Observation e9fa5e75-9146-4be3-b3ce-2a4d7c1b470c · inbound

Efficient Video-to-Audio Generation via Multiple Foundation Models Mapper cites this paper.

Efficient Video-to-Audio Generation via Multiple Foundation Models Mapper FoleyGen: Visually-Guided Audio Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T05:50:17.649506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:50:17.649506Z digest=sha256:e3c8ae7b6af4bdad7a2a487e7d97d018f2cf99ea79eea535a33fc383e9bb9313