Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:33:14.254397Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2607.19859.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:33:14.254397Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 762d9abe-c9e1-45d6-b5c6-8b016af7e3d8 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e95087c6-65a3-4578-b689-5d7c7a80761a · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba334b15-2f8f-4727-adbf-68c4606b6ace · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 032aedcf-b139-4bfc-bec3-cb8564c8078b · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis SoundStorm: Efficient Parallel Audio Generation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60ab2a50-c696-4b5d-905b-9e54efd22029 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb22d84c-56d5-40f1-a14d-1a95d0c6e2be · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis E2 TTS: Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 559dbbb9-104c-4d86-b497-679ab7fe43a6 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c255864a-fb98-409a-ad4e-539c5adf2442 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ad4210d-96e1-4a88-bc55-07a6f53cb46b · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis Fastspeech: Fast, robust and controllable text to speech,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0d3e1bf-75f0-4f90-b74b-e46458c72f28 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis FastSpeech 2: Fast and High-Quality End-to-End Text to Speech
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 927b2462-ee40-49c2-987a-e358666ffbf7 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis V oicebox: Text-guided multilingual universal speech generation at scale,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feb60fb7-d20c-49c9-93c2-cc2bafa5bac3 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0a25f98-c10c-4f9c-8a48-09af0406945a · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis LLaMA: Open and Efficient Foundation Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b382ecb4-1ab2-4503-b3d6-547b06d9e204 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2229f2bd-e850-4bf5-b9e1-0b068a98df86 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis Neural discrete representation learning,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fc1f77a-fc24-40d7-9683-4155eee4a389 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis Soundstream: An end-to-end neural audio codec,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dabd932-87a8-4315-9da8-06b623d3b42a · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis High Fidelity Neural Audio Compression
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c54927a-3390-40db-af24-1c7d8fa7252b · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis Hubert: Self-supervised speech repre- sentation learning by masked prediction of hidden units,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddee6882-f116-4dbf-ad7e-24097741fbee · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis Wavlm: Large-scale self- supervised pretraining for full stack speech processing,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98219372-94ca-4c97-b7cd-f74e065db327 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis W2v-bert: Combining contrastive learning and masked language modeling for self-supervised speech pre-training,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 505248d2-3a49-4223-b368-65e25321864d · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcfe1bb0-7916-43a6-ad3b-06d4a79bf3d3 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c9468fe-a430-472d-8256-dce62ab51a78 · outbound
StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.