Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T16:01:52.889397Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 3 inbound Pith citation observations for arXiv:2508.19098.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T16:01:52.889397Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-27T21:18:22.911332Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T19:47:19.713205Z
69 of 69 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 958bcb74-b0d2-4b6d-b750-76a12a654377 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce96466a-3f3c-498a-b4c6-699d1b217b7a · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing, May 2022
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee274c21-6231-4378-b2c4-d7d16bd2c174 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Rethinking lossy compression: The rate-distortion-perception tradeoff
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9568581b-d05c-427d-bcda-b47528a0d1a1 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Audiolm: a language modeling approach to audio generation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 90aa4c5a-2bd8-46eb-8908-16a6a917b2c7 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models, April 2025
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 19a81d71-38a8-4e58-8a48-58b6b09987ee · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Wavlm: Large-scale self-supervised pre- training for full stack speech processing
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 53ce4747-da48-485a-b0f5-1a133c8439b7 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis VALL-E 2: Neural Codec Language Models are Human Parity Zero-Shot Text to Speech Synthesizers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 745110e8-993f-4339-8a4b-51e03547cfbb · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57fdbef0-4a32-4f96-8ff3-dfb48569668b · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis High Fidelity Neural Audio Compression, October 2022
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 14557648-9203-4b77-94aa-6dde161d30c6 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5adda8a6-57e9-4477-82d1-a3a07bf6a619 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b85f4d6d-3bfa-4a61-a653-8b689b6c30b2 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis E2 TTS: Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67b6aa19-bfba-4ecf-9d2b-7792931a7eb6 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Taming Transformers for High-Resolution Image Synthesis
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ae6eafc0-1ccf-4c0c-b577-b0739db50a6e · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb7fc861-e8fa-4c79-a6fd-53f57b73239f · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Fast Timing-Conditioned Latent Audio Diffusion
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44e1c1ad-4a1d-4d7f-8636-d23a5b4e962c · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis E3 TTS: Easy End-to-End Diffusion-Based Text To Speech
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0c4c34e-f724-4f59-97a1-ce13369336bd · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 04207df1-0bfa-4b6d-9458-20207a08e057 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb6b0037-dbd3-4354-9c35-d3fcd46dd6fd · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis VALL-E R: Robust and Efficient Zero-Shot Text-to-Speech Synthesis via Monotonic Alignment
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e65ad875-0b89-4e36-a11c-da80d888d4e1 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Classifier-Free Diffusion Guidance
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aef40a6e-b847-4b40-8e4c-da3022bfe9f7 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Straightening out the straight- through estimator: Overcoming optimization challenges in vector quantized networks
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7833edb7-8922-4c02-aa15-c5cf3f7ac512 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Ditar: Diffusion transformer au- toregressive modeling for speech generation, 2025
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2e5b129c-c9c6-4923-bcd8-3039bb2510ee · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Mega-TTS 2: Boosting Prompting Mechanisms for Zero-Shot Speech Synthesis, April 2024
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 31454201-0a0c-44ca-837f-36efe3eea5f1 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models, April 2024
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 84602fb7-7bf0-4493-9b41-12650abc3872 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Libriheavy: a 50,000 hours asr corpus with punctuation casing and context,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fa7fcdce-99c5-42f5-a527-58038e72b5a1 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Analyzing and Improving the Training Dynamics of Diffusion Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bc03816e-8704-4a65-a3bf-6d7772f6b3be · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e0749a-329f-4bde-968a-37b2e22944bd · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis High-Fidelity Audio Compression with Improved RVQGAN
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5cfa107d-b223-498f-870b-eec2eb05cec5 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis BASE TTS: Lessons from building a billion- parameter Text-to-Speech model on 100K hours of data, February 2024
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d5de89ad-5244-4d9f-98e4-a6e6c7f4439d · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis V oicebox: Text-guided multilin- gual universal speech generation at scale
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eebed74a-9958-470e-95d0-23b257fd04be · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis REPA-E: Unlocking V AE for End-to-End Tuning with Latent Diffusion Transformers, April 2025
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 91a25cc2-7bd8-42d5-8329-6c8aba46c0a8 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Neural Speech Synthesis with Transformer Network
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e8c5c59-9594-4e39-b972-1cd22596e5c6 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Autoregressive Image Generation without Vector Quantization
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54dc638d-e939-4cb4-8af7-14894dfa04f4 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fbfafa0-f539-453a-95f2-15b3161ffe4c · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 439fbd8b-bb4f-46e5-95be-fc4997f507ad · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis LibriSpeech-PC: Benchmark for Evaluation of Punctuation and Capitalization Capabilities of End-to-End ASR Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 95e71bba-b9f8-4aa8-9666-b8f109020006 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Autoregressive Speech Synthesis without Vector Quantization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40e3b997-26db-4c2a-bd81-b5ea98b4355a · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Finite Scalar Quantization: VQ-V AE Made Simple, October 2023
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ca2f0a00-cd01-453c-ab18-6e562a1fbf90 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis HALL-E: Hierarchical Neural Codec Language Model for Minute-Long Zero-Shot Text-to-Speech Synthesis
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c9ac590-e942-4621-8e3c-34b488390f90 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Librispeech: An ASR corpus based on public domain audio books
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 25b5268c-6ecf-4fe3-9b77-07cb518ff5bd · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Robust speech recognition via large-scale weak supervision
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e2ca104-69e1-476a-8e25-097231b26321 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Revisiting over-smoothness in text to speech
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d78640a-6f08-41fe-8d3a-5d48bcd6aad6 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7183e728-81f3-4342-807a-9d614121e605 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis AutoClip: Adaptive Gradient Clipping for Source Separation Networks, July 2020
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5943c9b-07cd-4172-98a2-2b6298578e77 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis GLU Variants Improve Transformer
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c67adb7-6ae8-4540-9c83-0d94bb0ab2bd · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers, May 2023
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aeb04ce9-c2aa-4a98-9784-484353023d73 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7350a1b-82de-4554-bdc7-c4b14864f0d5 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Ella-v: Stable neural codec language modeling with alignment-guided sequence reordering
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8d0864e8-9d61-4847-b4ae-8007d02a4b4a · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Steinmetz, Jordi Pons, Santiago Pascual, and Joan Serrà
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 88ac9b82-424e-4284-90e6-7e3474890267 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis RoFormer: Enhanced Transformer with Rotary Position Embedding
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e22d374-289f-4f3a-9e3a-fd63b7c1ac92 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis NaturalSpeech: End-to-End Text-to-Speech Synthesis With Human-Level Quality
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1d9c93ff-5827-44e5-a4de-57432d2981a0 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Continuous Speech Synthesis using per-token Latent Diffusion, October 2024
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e0e7da8-4df0-4875-9f8c-592ba1c12b98 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Neural Discrete Representation Learning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7a870c0-8ae9-4cb2-b7d2-6361397113fe · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5b0d8c7-9cc3-4359-8b41-ce886d522b74 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61f5c358-f70e-48f2-9779-a8badee88a3c · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d998edd-d252-4ed5-a46d-6df9e126adaa · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer, October 2024
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 536b2ec0-62b9-481b-b8ee-ab4ccc1042b9 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Towards audio language modeling -- an overview
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 981dec4e-3a18-4f0b-a20a-d12e616454d4 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis RALL-E: Robust Codec Language Modeling with Chain-of-Thought Prompting for Text-to-Speech Synthesis
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa13fcd6-7014-4990-9911-08281a879493 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis On Layer Normalization in the Transformer Architecture
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02ee09ec-e051-4671-805d-6f91e1e6141b · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Reconstruction vs
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cc1ff55e-0c36-4212-9fcd-d9cfb7916350 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffefd789-2cff-4087-abec-ee7e034b4dd6 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think, February 2025
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0fb202ed-c747-45d3-918b-8fe81290ee9f · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Soundstream: An end-to-end neural audio codec
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9ac53ff-9952-42b0-9262-3634970b2ed3 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Weiss, Ye Jia, Zhifeng Chen, and Yonghui Wu
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 449586ca-366b-4904-8afd-99a0efc17c80 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Root Mean Square Layer Normalization
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9c9843c-e0b0-4d37-9727-f0540ccd7d32 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c846336-a41f-4239-85a9-dc8976d02ace · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis ,𝑁𝑆𝐵,𝐶!,𝑁 + UpsamplingBlock Channel toSpace ChannelDuplicating𝐵,𝐶
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c35e1bbd-d098-48b1-b4de-3d045c067d50 · outbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f877a92c-605b-4346-bb9a-935b28c1d987 · inbound
On the Distillation Loss Functions of Speech VAE for Unified Reconstruction, Understanding, and Generation CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c6001f44-543d-4894-8ba1-62ff1c34b900 · inbound
SemaVoice: Semantic-Aware Continuous Autoregressive Speech Synthesis CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6947956a-a27e-4039-9221-c3793b937b2b · inbound
VoxCPM2 Technical Report CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.