Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T11:52:50.598080Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2607.10421.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T11:52:50.598080Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9b3fc724-f212-4470-a95c-1e1e5c606eda · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Diffused responsibility: Analyzing the energy consumption of gen- erative text-to-audio diffusion models,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6a288de-c2f6-4954-a82e-4dbfac3e3d92 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Denoising diffusion probabilistic models,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c342c0a4-bc17-4eea-be6b-f55cbbc08972 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Flow matching for generative modeling,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b8edfd2-2c7b-48ec-90d0-4d4028f8f01b · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Audioldm 2: Learning holistic audio generation with self-supervised pretraining,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1630e7b-902a-4ded-885f-20542620c3c8 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Make-An-Audio 2: Temporal-Enhanced Text-to-Audio Generation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f8bcb6a-d5b0-4c67-809a-332b91dc07f7 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5225392a-18c1-4d74-bb3a-67134f339fb3 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2b5beb3-0265-45b8-bdae-30cdb9109561 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Impact: Iterative mask- based parallel decoding for text-to-audio generation with diffusion mod- eling,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32b5affa-9fcf-4104-9966-5aad433aeb18 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Generative audio language modeling with continuous-valued tokens and masked next- token prediction,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91139c4a-cb0e-4156-86fc-d0997f70839c · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 787e77e3-43c1-41c2-9d41-7ce15f754e33 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Resonate: Reinforcing text-to-audio generation via online feedback from large audio language models,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6564cabb-cb6a-43b4-95e3-db7286f4326d · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Consisten- cyTTA: Accelerating Diffusion-Based Text-to-Audio Generation with Consistency Distillation,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 448a95b5-f72f-46fa-8bb6-02ad2c8b0be9 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation AudioLCM: Text-to-Audio Generation with Latent Consistency Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4836bf38-4f4a-4b3e-974a-49ec1ab4827e · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation AudioTurbo: Fast Text-to-Audio Generation with Rectified Diffusion
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a0e3136-f195-497d-9cde-37a243fb63c7 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Soundctm: Uniting score-based and consistency models for text-to-sound generation,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79a2d7b3-b564-4331-a444-46dfa7549aa6 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Fast Text-to-Audio Generation with One-Step Sampling via Energy-Scoring and Auxiliary Contextual Representation Distillation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84884172-5761-4be8-94fb-e307d06c3304 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Meanaudio: Fast and faithful text-to-audio generation with mean flows,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ce94d7e-4383-4606-a290-c826b768d528 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Representation Fr\'echet Loss for Visual Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41936e17-7ad3-4898-8b0f-5829edb40ab1 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation One-step Latent-free Image Generation with Pixel Mean Flows
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cf87a7e-7209-4421-9242-379dbae502c2 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Improved mean flows: On the challenges of fastforward generative models,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cb6e32e-3213-4e95-bfd7-79fd88cbda9f · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Rethinking the inception architecture for computer vision,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5389094d-485c-43dd-9e7b-e3beda2007f8 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation A convnet for the 2020s,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc496b24-9abc-41da-b947-1af532dd2de9 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Masked au- toencoders are scalable vision learners,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14000e90-98c1-4845-8fe8-becba052302f · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation DINOv2: Learning Robust Visual Features without Supervision
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1df227df-61fa-433e-918f-8991112aeb6c · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a6dce11-03d4-493f-a423-7b12a6b009b4 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Learning transferable visual models from natural language supervision,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80f65fbd-d9d9-438b-9f20-c2fceea7d0c6 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Panns: Large-scale pretrained audio neural networks for audio pattern recognition,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f25018d-04b1-4c8d-957b-3f56b043c333 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Efficient Training of Audio Transformers with Patchout
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35666e69-0923-4e17-9b60-173abe7283e8 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation BEATs: Audio Pre-Training with Acoustic Tokenizers
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5445dba3-c174-45bb-ba8e-1d0b7b374042 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Masked autoencoders that listen,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2adef9c-7570-46fb-80b7-f5cff6c37478 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Mean flows for one- step generative modeling,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ad7b97e-fa98-496e-a94b-9ecddec4c3bd · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Consistency models,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7f7f1ae-7dde-48a0-9dc6-f6afcaa0ef75 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Classifier-Free Diffusion Guidance
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5c39f40-ce4e-49fa-ab65-4b6b2850d562 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Cnn architectures for large-scale audio classification,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31071348-ec9d-4de3-ab60-b9d04e90aa26 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Fr ´echet Audio Distance: A Reference-Free Metric for Evaluating Music Enhancement Algorithms,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f1c48ba-5e77-4be8-a438-f33b33454261 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Audiocaps: Generating captions for audios in the wild,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f456aa68-c034-4fe6-8362-323bf9306d49 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Wavcaps: A chatgpt-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cefcfe44-49ef-498a-89d2-a368fb3c3f47 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Improved techniques for training gans,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e25f93f-cb18-4481-9ff8-7f32405e7f25 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeee2452-af75-4daf-80e5-0033f521df71 · outbound
FdAudio: MeanFlow-Anchored Fr\'echet-Distance Post-Training for One-Step Text-to-Audio Generation Scaling instruction-finetuned language models,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.