Pith. sign in

Paper Citation Record · LEDGER

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation

As of 15 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2607.29148.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.29148 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T12:47:00.255183Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b7faae89-7517-47d2-8fbd-566671c26b64 · outbound

This paper cites Foleygan: Visu- ally guided generative adversarial network-based syn- chronous sound generation in silent videos,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Foleygan: Visu- ally guided generative adversarial network-based syn- chronous sound generation in silent videos,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.420921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.420921Z digest=sha256:29d036b98bab170923c18411ed227166820cfda649e9f62c7ae33e09d81915c8

Observation e2e21079-58d3-4071-840c-8f226764a765 · outbound

This paper cites The x-lance system for dcase2023 challenge task 7: Foley sound synthesis track b,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation The x-lance system for dcase2023 challenge task 7: Foley sound synthesis track b,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.471416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.471416Z digest=sha256:b2f96b26c9b59fa2e42881be25d55bd318861ec21f9a36c9f360bcdfcc2bb662

Observation 44e02b7d-27d6-494a-adef-748b392c51f2 · outbound

This paper cites Mtdiffusion: Multi- task diffusion model with dual-unet for foley sound generation,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Mtdiffusion: Multi- task diffusion model with dual-unet for foley sound generation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.566213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.566213Z digest=sha256:2cad88c061192a918469193b6f145f4a278192362fac4fe1f49a21a0f2487f8f

Observation a8cc5034-618d-477f-bfa9-f77711cedb6a · outbound

This paper cites Rhythmic foley: A framework for seamless audio-visual alignment in video-to-audio synthesis,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Rhythmic foley: A framework for seamless audio-visual alignment in video-to-audio synthesis,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.673463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.673463Z digest=sha256:074fe0afa2dcdff3e04ce9fe6c7ccf2f4b4d4d33b1c92869b026526e80380745

Observation 8b3ef8bd-5fbc-4dbc-8fc3-b9955bbdad09 · outbound

This paper cites T-foley: A controllable waveform-domain diffusion model for temporal-event- guided foley sound synthesis,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation T-foley: A controllable waveform-domain diffusion model for temporal-event- guided foley sound synthesis,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.781442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.781442Z digest=sha256:0c0271811a7fdc881a5eae78cfae3e8a1502bc7aacb6e5781ca2d66095c478f2

Observation 66d2d469-9ce7-4885-ada3-0d6d8c8eec1f · outbound

This paper cites Text-Driven Foley Sound Generation With Latent Diffusion Model.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Text-Driven Foley Sound Generation With Latent Diffusion Model

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.862868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.862868Z digest=sha256:db4516fc2a2062c88e8325dcbb9ddc5a5cbe64734f937c84778c1d7e69bbd67b

Observation 8e6509e6-dac5-48c8-9391-0c7876ebc211 · outbound

This paper cites AudioLDM: Text-to-Audio Generation with Latent Diffusion Models.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:58.947785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:58.947785Z digest=sha256:0ad8e9cef62ed217ec534c221cf3c789b3b0d642a12d04304a18f95ce89ed33c

Observation 13aa4e11-f584-4fb5-a94b-3f1d056adb00 · outbound

This paper cites Audioldm 2: Learning holistic audio generation with self-supervised pretraining,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Audioldm 2: Learning holistic audio generation with self-supervised pretraining,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.077301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.077301Z digest=sha256:80a0e86134c769f2bc48f23a3dea04e7b2091160d6451e4d6525e15d7e5494d6

Observation 5fcd8b7b-cf3a-46e5-8093-3caeff0681c3 · outbound

This paper cites Generative speech foundation model pre- training for high-quality speech extraction and restora- tion,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Generative speech foundation model pre- training for high-quality speech extraction and restora- tion,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.169116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.169116Z digest=sha256:73fb6177c31ae105b70adac923352d1f5b25b0280efdbf87d6ad11fd693fa443

Observation fc5404f4-d96a-4b8e-8479-6aa933bd6c08 · outbound

This paper cites Dual-path rnn: Ef- ficient long sequence modeling for time-domain single- channel speech separation,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Dual-path rnn: Ef- ficient long sequence modeling for time-domain single- channel speech separation,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.293363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.293363Z digest=sha256:9de7a8a413dfb6f2b0e83ffb036992d0675b09f89983915bec1c0c3e5bda6f3f

Observation 6346c9a8-642c-4cbc-870a-67acc081768c · outbound

This paper cites A study on speech enhancement based on diffusion probabilistic model,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation A study on speech enhancement based on diffusion probabilistic model,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.406840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.406840Z digest=sha256:03806a78d2758ee5388ec1e053a1befa076a919e8199f0aa132076034c621985

Observation f42dd0bd-c244-476c-b2fe-43aa44a00dcf · outbound

This paper cites Unsupervised single-channel audio sep- aration with diffusion source priors,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Unsupervised single-channel audio sep- aration with diffusion source priors,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.456317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.456317Z digest=sha256:a159cd66333faf00772a89b0095fe7973c2c6c57a352b728032cc0d3c4a625ac

Observation 367a5819-130d-4faa-9a77-92d11bd494ac · outbound

This paper cites Mambafoley: Foley sound generation using selective state-space models,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Mambafoley: Foley sound generation using selective state-space models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.525316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.525316Z digest=sha256:0a09924ad6d84db3e344d6a7b3a1e7e4435bc031106f27c14b4cc62816d8a7aa

Observation f6f8ba3a-c65d-476a-b286-830334482bc2 · outbound

This paper cites Full-band general audio synthesis with score- based diffusion,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Full-band general audio synthesis with score- based diffusion,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.615845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.615845Z digest=sha256:05b4f42a66f7e29775c63cc2999aeb8f92bb9cbf3ec31f65b4d44f916fdfc4f8

Observation cacf177a-1e36-4216-8978-fd6a2a216346 · outbound

This paper cites Diffwave: A versatile diffusion model for audio synthesis,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Diffwave: A versatile diffusion model for audio synthesis,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.663901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.663901Z digest=sha256:688ef58e0a2b5802dafeb328ed489825f5f937cca6909afd28d0c4c96b6d64f4

Observation 14582748-2ac1-4ebd-9b42-f39e00e5c434 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.714092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.714092Z digest=sha256:3e6a57b102a80281762aa33bd845dea3cffea78333f2f2cf089c08b40538f38c

Observation 56df86ed-dd3c-467a-a138-6a7803d005f7 · outbound

This paper cites Fregrad: Lightweight and fast frequency-aware diffusion vocoder,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Fregrad: Lightweight and fast frequency-aware diffusion vocoder,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.773065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.773065Z digest=sha256:0c973581ed9f68d0cf452e02905206ed3780b6af57a31566b07168081b5b2d2b

Observation 2eea032b-d09e-4b00-a289-3660bdc6e82c · outbound

This paper cites UnDiff: Unsupervised Voice Restoration with Unconditional Diffusion Model.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation UnDiff: Unsupervised Voice Restoration with Unconditional Diffusion Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.820387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.820387Z digest=sha256:cee5defdfa0d07caaa43b9f061a9cf8a07c247a0e621adb1d169f43ac4966293

Observation 690cc6c0-2e86-4641-8332-34544f01c746 · outbound

This paper cites Solving audio inverse problems with a diffusion model,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Solving audio inverse problems with a diffusion model,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.906532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.906532Z digest=sha256:236f166544170b75fcab4d20073a71590176d29359913a903f4428028104be49

Observation 65e250d6-84ca-4ec1-b5c3-87298b4826ae · outbound

This paper cites Foley Sound Synthesis at the DCASE 2023 Challenge.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Foley Sound Synthesis at the DCASE 2023 Challenge

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T12:46:59.967862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:46:59.967862Z digest=sha256:367875a83dda0fd210fecf9fe8776904c74a110244cd85c3a6c72fb3cb0df641

Observation c9b1bd9d-29cc-4567-9f64-681b9d0ba6b2 · outbound

This paper cites Scalable diffusion models with transformers,.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Scalable diffusion models with transformers,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T12:47:00.032718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:47:00.032718Z digest=sha256:89ee54a41e1093dd20c7e9ba177c09ea1ba0c702de5ab89f8842635cdf048886

Observation b3d4713e-705c-430f-abac-c7374293af9e · outbound

This paper cites General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T12:47:00.111103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:47:00.111103Z digest=sha256:986acee89d24698b7f9d958802f266da251e1456a0690e99988006d7a33acd5b

Observation cf6d8bac-7564-4a98-8948-0629a54414e6 · outbound

This paper cites Decoupled Weight Decay Regularization.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Decoupled Weight Decay Regularization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T12:47:00.165216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:47:00.165216Z digest=sha256:fb44631d28a5b872e7af98a6b6f04ba3f4892272ac169e0abd523067ce469bf8

Observation b26af08f-57c4-46ea-8b12-a2f75bbfd77a · outbound

This paper cites Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms.

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T12:47:00.255183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:47:00.255183Z digest=sha256:ee19b379da0065a777025f19962b07a3b65962aaca93810a3f848c22d87f4441

Pith citing papers

No inbound Pith citation observations are available.