Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2502.18924.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:30:34.140760Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T20:50:11.241591Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation cdbe0ec9-7038-4163-a31d-7bc3d19ebd5f · inbound
TCSinger 2: Customizable Multilingual Zero-shot Singing Voice Synthesis MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 587953c3-962d-4b96-8ae5-507bfbef7791 · inbound
ACE-Step: A Step Towards Music Generation Foundation Model MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 099d9eb2-9a58-43c6-b9b0-a7be49a96e24 · inbound
Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6bd46b6-78ae-4182-87a8-0c239981490a · inbound
JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 03cecab6-b8f0-4506-8765-6f7727eddc1e · inbound
The Thin Line Between Comprehension and Persuasion in LLMs MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a800ec70-a113-4f3c-8d23-8a6a1ba1b0f0 · inbound
Exploiting Leaderboards for Large-Scale Distribution of Malicious Models MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 693a491d-ead1-4089-a240-50699ac46ec3 · inbound
JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c783b992-901b-485c-ae29-8d885f5d5cd6 · inbound
Attention2Probability: Attention-Driven Terminology Probability Estimation for Robust Speech-to-Text System MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dd14df9-6bc9-44e4-a47a-e65792cea62a · inbound
DiTReducio: A Training-Free Acceleration for DiT-Based TTS via Progressive Calibration MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4ed716a-13e0-4bc3-a677-661395c9a330 · inbound
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9625e5cf-2220-4ca0-8e39-cb1d27fd56b5 · inbound
OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1662e086-ac3a-4e49-a673-016f13136596 · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4c39120-64ca-43b2-befd-ce6d88be8442 · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04768d21-8e5a-4997-9c49-9a26dedad9ca · inbound
CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4099f967-69ac-4231-b82a-512ba525982c · inbound
From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2590a872-a3f4-425c-86d7-f6e915558813 · inbound
RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3ad9fbba-a588-4172-919f-eb6559d93ab0 · inbound
RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f700083f-6c0b-4c7d-afe0-7f7c12ce8eba · inbound
SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 768e4f15-4a57-45e2-adb1-22e3139c3de8 · inbound
WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 111fa7d6-a449-4200-a13e-c0646cbd6737 · inbound
VoxCPM2 Technical Report MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d62cfc91-7398-45ee-acda-3dc9648f1de1 · inbound
dots.tts Technical Report MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6ee9bf96-8091-4db3-a043-d2fa1fc3cf9b · inbound
Reliable Neural-Codec Text-to-Speech by ASR Self-Verification and Distillation: Near-Zero Catastrophic Failures Across Models and Codecs MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2bcc5c2b-2ddf-4239-a728-12ede820e97e · inbound
Reference-Driven Multi-Speaker Audio Scene Generation from In-the-Wild Priors MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9ab8c22f-96c7-43be-8eaa-9bd5e66b0d74 · inbound
EmoInstruct-TTS: Dual-Path Instruction-Guided Emotional Speech Synthesis MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 910e558f-fe07-4e39-b823-61df4e107646 · inbound
Joint Residual Reweighting for Classifier Free Guidance in Flow-Matching Zero-Shot TTS MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 763bbe90-8755-42a1-9e6e-4e16b0463491 · inbound
Joint Residual Reweighting for Classifier Free Guidance in Flow-Matching Zero-Shot TTS MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4d72292d-bc6e-491b-a2d7-affa35327ce3 · inbound
DETECT-3B-Omni is Agnostic of Content and Demographics MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 966fa3d7-be6a-434e-9f6b-365d6f941014 · inbound
ReGen: Hierarchical Multi-Prompt Representation Generation for Efficient Waveform Diffusion Models MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1562ad06-8805-48b2-a8ad-4d7b16698ae6 · inbound
DLLM-TTS: Block Discrete Diffusion Language Model for Text-to-Speech Synthesis MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f8c3cba-3765-47a0-b2de-46647600cdaf · inbound
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf3d068e-ceb0-4675-abbf-063188f005fa · inbound
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.