Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2409.03283.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:26:19.124096Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T08:29:41.338585Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d20dfc8c-8fda-41d8-81f3-2483d87ab40e · inbound
F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e3754d63-1370-4fd4-8898-c1ddcb4dc393 · inbound
CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab9f82f4-1ab7-4ad4-82d3-3aab0706d44f · inbound
Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ba236b4-77cf-483f-80a7-d4709efe4c0d · inbound
Accelerating Autoregressive Speech Synthesis Inference With Speech Speculative Decoding FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c380d880-2de5-43e2-b763-1b8e6964a994 · inbound
CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 235bff6a-1867-49b6-87de-bd2096abb91f · inbound
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 594d5d61-61ff-4221-becd-05815f436794 · inbound
Accelerating Diffusion-based Text-to-Speech Model Training with Dual Modality Alignment FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7425d4dc-2ea1-4ce5-9d9b-c5eae321ef72 · inbound
AudioGenie: A Training-Free Multi-Agent Framework for Diverse Multimodality-to-Multiaudio Generation FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47614f9a-f3b0-4357-b723-32f725028c9f · inbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f04ef65d-a581-47f3-8108-7d8574e282b2 · inbound
IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4cb8f56-6738-412c-b666-4824b108b49c · inbound
JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 30f9b39f-def3-4812-afda-b754c9957879 · inbound
Speech Quality Assessment Model Based on Mixture of Experts: System-Level Performance Enhancement and Utterance-Level Challenge Analysis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d117a21e-b6f3-46f3-b34f-e7ac10819055 · inbound
ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f720103-5c55-497f-8569-2b43f5e60094 · inbound
NonverbalTTS: A Public English Corpus of Text-Aligned Nonverbal Vocalizations with Emotion Annotations for Text-to-Speech FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd611872-f44b-4e2d-aa23-8a946cd85e5e · inbound
SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c671d657-7cf7-46fc-a3ff-b7c02cdf8e9e · inbound
NVSpeech: An Integrated and Scalable Pipeline for Human-Like Speech Modeling with Paralinguistic Vocalizations FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04207df1-0bfa-4b6d-9458-20207a08e057 · inbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cefde6a1-a184-456f-8e52-c5e09ce4e093 · inbound
FireRedTTS-2: Towards Long Conversational Speech Generation for Podcast and Chatbot FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfafd71d-d7a0-4111-ac51-e2955428a5c4 · inbound
FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d28099c9-99dc-49a7-8377-68848241ab74 · inbound
Ovi: Twin Backbone Cross-Modal Fusion for Audio-Video Generation FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e26d98a2-a154-49b9-9b21-89d5298ec760 · inbound
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 823a16f6-e458-4820-adef-e45758a0ca7b · inbound
EchoFake: A Replay-Aware Dataset for Practical Speech Deepfake Detection FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 77153097-b6dd-4302-877f-27f9e3db101b · inbound
OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dff7dc9f-38a9-4243-a83e-443a16023c49 · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 46d95095-df79-4898-95cf-9d0b808f6a00 · inbound
AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be6446a3-77f2-498c-818d-b1dbf5582866 · inbound
SemaVoice: Semantic-Aware Continuous Autoregressive Speech Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee013815-f54d-44c8-92bb-fac7cfed4656 · inbound
SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5554fe46-4731-44ee-8b2e-3cec4b45bde0 · inbound
UniVocal: Unified Speech-Singing Code-Switching Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4476e729-e5e4-4ced-aa61-c1d5762c8250 · inbound
WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d6c9f23d-4232-4d1b-b5be-cbde1fa35f93 · inbound
VoxCPM2 Technical Report FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 45724a0f-9436-45d4-983c-cf0fbc79f786 · inbound
TLDR: Compressing Audio Tokens for Efficient Autoregressive Text-to-Speech FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f3a401c-6ffa-45d9-8b66-5b3376effd5b · inbound
HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b4dfecfa-5f66-46ce-9ec4-c8d33a0bfd67 · inbound
FlashTTS: Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 58c08025-067f-41e3-96ba-0ad7c52fcac6 · inbound
End-to-End Training for Discrete Token LLM based TTS System FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 24829c5b-27c9-47d4-acbd-828dfe98d711 · inbound
One-Step Token-to-Waveform Generation with MeanFlow in Latent Space FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0479e9c8-ef06-4c56-97af-b5a27e608693 · inbound
Streaming T5-based Text-to-Speech Synthesis with Limited Lookahead FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d9b06688-7d9e-4526-84ae-34abc4897273 · inbound
FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 191
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a0d0e77a-334f-45cf-a8b6-7686f25d5dba · inbound
Enhancing Flow Matching with A Unified Guidance Framework for Efficient and Robust Speech Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7916e613-da85-4173-8696-8df411191f06 · inbound
ReGen: Hierarchical Multi-Prompt Representation Generation for Efficient Waveform Diffusion Models FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c9468fe-a430-472d-8256-dce62ab51a78 · inbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1eae712-ec60-4c69-95c0-911f42e4aa10 · inbound
Faster IndexTTS-2: Accelerating and Streaming Autoregressive Zero-Shot Text-to-Speech Synthesis on GPUs FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 680b1adf-f1f3-48ed-9db1-03bf0ed52b91 · inbound
Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.