Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T20:31:24.494060Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 3 inbound Pith citation observations for arXiv:2409.18512.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T20:31:24.494060Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T15:21:28.809697Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-03T00:47:30.552861Z
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ad99742c-3ffc-4e72-bef6-190a5430f71d · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Improving language understanding by generative pre- training
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 30f9dfec-c267-459c-b604-d7a7e2cf491e · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 41a2b1dd-079b-45c7-b714-8916ac2a2a2d · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Speak, read and prompt: High-fidelity text-to-speech with minimal supervision
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 037ffe26-1e20-43f8-b9e4-ea3e935ba457 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Hubert: Self-supervised speech representation learning by masked prediction of hidden units
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f01f4e79-4c5e-45b6-bed7-27dfaa3c9b4c · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Learn- ing speech representation from contrastive token-acoustic pretraining
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 25331a38-e140-4b57-858c-6ff85959fe2c · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS High Fidelity Neural Audio Compression
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ce977247-b194-45cc-aab6-fd1a9270f528 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 845d86b3-8597-459f-b362-021156237d04 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS AudioPaLM: A Large Language Model That Can Speak and Listen
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b5c32be0-0f97-46bb-b4c3-b1ecb7d43a1d · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS VioLA: Conditional language models for speech recognition, synthesis, and translation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f44803bf-8ca9-4ce2-99ac-e5a902bf5f66 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Large lan- guage models are zero-shot reasoners
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8b51d302-f83c-4283-b395-0dec0540f6d8 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Scaling instruction-finetuned language models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d5625e6c-f2a4-4c3b-8f7c-1c155660c07b · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Retrieval-based prompt se- lection for code-related few-shot learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8e68ddf2-4fd1-445b-880d-4961b4cf0ce5 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Learning to prompt for continual learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 19f305a6-b778-453d-a025-e067509cd618 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Automatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d8ba8b3f-8e7a-445f-b252-fe83809d15d0 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Universal information extraction as unified semantic matching
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ed3b68eb-c4ac-4634-99ea-8eb79382650d · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Sentence-bert: Sentence embeddings using siamese bert-networks
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b83fb258-e51b-4a4c-8554-331dbcdcf692 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Controlling Emotion in Text-to-Speech with Natural Language Prompts
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 041024c9-5f47-488e-9eb4-24126add3958 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3be1f78d-e8b2-4032-8a1d-270876184055 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3d1c4d59-b651-4169-8a74-dc93eb7c4d49 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Zmm-tts: Zero-shot multilingual and multi- speaker speech synthesis conditioned on self-supervised discrete speech representations
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d833eacc-5cb3-480c-84da-9624a9678e28 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9e2f48f9-a5c9-4394-aa23-73f9e3ba91ee · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS GPT-4 Technical Report
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ee69e7e0-17f7-4555-b2e8-474e91aa36a8 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Intonation and emotion: influence of pitch levels and contour type on creating emotions
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 113cc4ef-a5d3-44af-8655-3f51c634c89c · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Analysis of emotionally salient aspects of fundamental frequency for emotion detection
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c28df821-9ee5-4f54-800d-10863e76941a · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Pitch in emotional speech and emotional speech recognition using pitch frequency
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9efe4297-055b-4a9f-89ab-2dacc6fb56ca · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Communicating emotion: The role of prosodic features
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 855181f3-0960-4072-91bf-c1a34991a917 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS The global k-means clustering algorithm
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 63f5efe2-bfd4-4afc-b7a7-6756df737bc0 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Generalized end-to-end loss for speaker verification
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1ee941bb-36ce-4200-8d2a-43f8bc2d17a1 · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Wavlm: Large-scale self-supervised pre- training for full stack speech processing
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8c1e784a-8609-4a47-b1bb-6dc4731fc21d · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS emotion2vec: Self-Supervised Pre-Training for Speech Emotion Representation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 005b0f11-cf72-4691-8592-be398e32260c · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Improv- ing prosody for cross-speaker style transfer by semi-supervised style extractor and hierarchical modeling in speech synthesis
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f89302da-3746-4e9b-b854-b0ed11be929d · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Minilm: Deep self-attention distillation for task-agnostic compression of pre- trained transformers
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 222b4700-42e0-42cf-a599-33585aa123cf · outbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Paraformer: Fast and Accurate Parallel Transformer for Non-autoregressive End-to-End Speech Recognition
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 04253d91-b32c-472f-8ddd-81d871d9c2c4 · inbound
The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 02415a3e-7d8b-468f-a09f-c4751d9e55d9 · inbound
The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 234b4815-02f5-4a2b-9e88-cd6e0ca0ddb4 · inbound
EmoInstruct-TTS: Dual-Path Instruction-Guided Emotional Speech Synthesis Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.