Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:21.978557Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2505.19931.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:21.978557Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:17.790140Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T14:09:22.202944Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 833af4bd-6ec8-4fc9-9f43-057dfcafb1ab · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b2d45a79-0e24-4778-83be-8dfdd4d75c84 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d908993a-a4d3-4a7b-aa87-c7b505f39f2d · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b29dd7fc-698b-458b-a56e-ed63be51e2c8 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Setup Baselines
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74270a73-3427-4686-ba11-8dfc8b9aa3e5 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bdc7e683-b8b1-422e-bfda-b94050802db0 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling U23B2018 and No
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 251ef0cc-0a20-44a7-bc65-7bb6071fe80c · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling V oicebox: Text-guided multilingual universal speech generation at scale,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41a2e567-a3f5-478a-b645-78cf8f326abf · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Neural codec language models are zero-shot text to speech synthesizers,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5fbba8c-c4de-4e24-80ce-2f7dffc1e352 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b4b6fda-367a-4857-a8dd-6ea3b5814775 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling V oiceCraft: Zero-shot speech editing and text-to-speech in the wild,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8d788c02-3a57-4d3a-9989-b5a4490215d9 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 684f42d8-25fd-44d4-bc51-97933e6dced9 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Autoregressive Speech Synthesis without Vector Quantization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6984a866-e405-425e-a30e-c7a29aa8b3a2 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f88baa60-07be-4a96-93d4-d4fd0ccd9cdc · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Flashspeech: Efficient zero-shot speech synthesis,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 371d11e3-772b-43c2-90e5-66f7d3129b07 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling NaturalSpeech 3: Zero-shot speech syn- thesis with factorized codec and diffusion models,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9305a402-9ae2-4fbf-9798-81d8a363a45f · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling E2 TTS: Embarrassingly easy fully non-autoregressive zero-shot tts,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ba44ab7-e2ce-4b97-b982-f296b1bfdde9 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41799ee3-71cf-4f7b-946d-f5e93b852970 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08b4c5b3-1170-45c1-a765-7c3edd8bbc23 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Naturalspeech 2: Latent diffusion models are natu- ral and zero-shot speech and singing synthesizers,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17c5c258-3e93-4b08-94f2-f983a1c4fd47 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Flow matching for generative modeling,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8aae7e9-da11-4c45-a8ee-211bf39df1fe · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling DMOSpeech: Direct Metric Optimization via Distilled Diffusion Model in Zero-Shot Speech Synthesis
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5bad944-e113-4a06-8a24-54f311e028d1 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab173b98-c8d8-45a0-b397-53ad21a74863 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Flow straight and fast: Learning to generate and transfer data with rectified flow,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e124aefa-74cf-4d60-82b1-0205d922c0a5 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling One-step diffusion with distribution matching distillation,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99ebb059-0973-4a04-ba54-396442ffb6dc · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Improved distribution matching distillation for fast image synthesis,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92312d31-4257-4686-8e20-50ce1b56e034 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling V oiceflow: Efficient text-to-speech with rectified flow matching,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2d5babc-5448-4bc1-812b-4f790f6bf6e7 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling All measurements are performed on single NVIDIA RTX 3090 GPU to ensure consistent hardware conditions
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 409fb9d5-cb23-48db-a668-1044ea7a94e1 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c17b967-96a2-4dab-ae22-72df9037102a · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Classifier-Free Diffusion Guidance
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b74a2ead-2055-4bb1-b258-19768e64983b · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Scalable diffusion models with transform- ers,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22b49046-9fab-4464-93d6-aa4854da86f7 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Convnext v2: Co-designing and scaling convnets with masked autoencoders,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63cf9ddb-25b8-49e3-bf24-99942b41f735 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Emilia: An extensive, multilingual, and diverse speech dataset for large-scale speech generation,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88557af0-f388-4c21-bd07-5ba4fa4ec765 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling V ocos: Closing the gap between time-domain and fourier-based neural vocoders for high-quality audio synthesis,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3bb9008-d1c6-4c8b-a2d7-d47575f68110 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling LibriSpeech-PC: Benchmark for evaluation of punctuation and capitalization capabilities of end-to-end asr mod- els,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34134a67-4d3d-4f83-8ef2-fd2e7af92e61 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Robust speech recognition via large-scale weak supervision,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90899022-b9f2-4eb0-b263-3c169974ec9f · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Paraformer: Fast and Accurate Parallel Transformer for Non-autoregressive End-to-End Speech Recognition
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a384c2e-ad52-43f6-a61c-cf2f639d9683 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Wavlm: Large-scale self- supervised pre-training for full stack speech processing,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 298d4cb5-81ce-4c41-a23d-7a84b1156450 · outbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling UTMOS: Utokyo-sarulab system for voicemos challenge 2022,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2d45a79-0e24-4778-83be-8dfdd4d75c84 · inbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.