Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2311.12454.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:32:24.075261Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T12:26:37.491138Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 1b46cae5-38b9-4e85-b7a8-66bdc50c31c2 · inbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ecc71c65-57fa-416a-ba47-2d1417464cdc · inbound
Towards Controllable Speech Synthesis in the Era of Large Language Models: A Systematic Survey HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06087c3f-9f4c-4b7e-a902-b7952bf232b7 · inbound
Low-Resource Text-to-Speech Synthesis Using Noise-Augmented Training of ForwardTacotron HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c73151f-6e9f-4e2d-b653-dddd8d5697c6 · inbound
Overview of the Amphion Toolkit (v0.2) HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ceaace19-f2da-4f7a-97fc-c90eb85c1b64 · inbound
Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e1a0fcb-be09-4d73-934a-3e235b15ca80 · inbound
Metis: A Foundation Speech Generation Model with Masked Generative Pre-training HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49e1554d-4460-4704-a07a-acb5915ffb54 · inbound
Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cbbd4e7-709a-402c-b2e3-a61fcb7b3d06 · inbound
Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 465a0c8c-bb1e-4e75-8021-eb5edcc90a2c · inbound
Towards Better Disentanglement in Non-Autoregressive Zero-Shot Expressive Voice Conversion HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 361f327e-a2bb-42b2-b314-ec4cf2d7f6ff · inbound
SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e7bddab-4b1c-421a-8be0-9ac862037162 · inbound
ClaritySpeech: Dementia Obfuscation in Speech HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.