Pith. sign in

Paper Citation Record · LEDGER

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis

As of 21 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2507.12015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.12015 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:00:35.341553Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:00:32.452371Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:00:35.528453Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a5be3642-8335-4aa2-bbcd-e0705651b4ea · outbound

This paper cites EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:00:35.612183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.452371Z digest=sha256:ae3d070b278150d6be37d9b3b32929eef93c9d272bdf8d19f2402e876b865e2c

Observation f3fd56f6-044e-4b73-883b-f093917cbfa0 · outbound

This paper cites Overview The overall architecture of EME-TTS is shown in Figure 1a.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Overview The overall architecture of EME-TTS is shown in Figure 1a

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:41.294741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.546382Z digest=sha256:f47635dac1268bac3c0cb66e129003e2533f0a7e03a0ab6f66a8ae71e323d1f3

Observation ac11db75-a129-483a-a4e5-ffa6d14a7314 · outbound

This paper cites an unresolved cited work.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:00:41.067248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.720040Z digest=sha256:13882402b19dcc254ff3d12e881fc85e1a68e364f3c7e9bfa70251883954effa

Observation e48b24f3-0364-4ccd-8c9a-fde12479c565 · outbound

This paper cites an unresolved cited work.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:00:40.949596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.761675Z digest=sha256:77a97bc350d7ae3179785bd1b7a0009f1cd7f660b344355466483dab4da57c00

Observation e559f109-d637-4a3a-879f-2e350a30d97a · outbound

This paper cites LQN25F020001), and in part by the Key R&D Program of Zhejiang (2025C01104).

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis LQN25F020001), and in part by the Key R&D Program of Zhejiang (2025C01104)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.797255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.787020Z digest=sha256:9ddeedf5609c707963489f7685f3cd70ed73fa88b1666d08f71d477eab6ac126

Observation 6a038c01-4ab0-470b-a154-038b0b632b2f · outbound

This paper cites Naturalspeech: End-to-end text-to- speech synthesis with human-level quality,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Naturalspeech: End-to-end text-to- speech synthesis with human-level quality,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.711282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.344834Z digest=sha256:76a8f0ce68ace16f47bf4eabbf19ffc3403ae581f0a53ae2ee6c6f6139e9bbd3

Observation 6b0b10a4-0295-460e-83e8-27b08c43b8e7 · outbound

This paper cites Attention is all you need,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Attention is all you need,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.623965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.844343Z digest=sha256:7ebc953e76a58c5beabd092a654c3d8ccdbfab8f868d4681cf68d8339f476afd

Observation 86a78097-fab6-45f8-8473-3cf0aa740f55 · outbound

This paper cites Fastspeech 2: Fast and high-quality end-to-end text to speech,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Fastspeech 2: Fast and high-quality end-to-end text to speech,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.424716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.970639Z digest=sha256:52b9f4b8e9013178b9d3ac7e96f86a3bc2f9cff0a8d983283918cc32ad5f225e

Observation 1bb5d809-8fe0-4f8a-a5d7-4eb3638ff385 · outbound

This paper cites Glow-tts: A genera- tive flow for text-to-speech via monotonic alignment search,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Glow-tts: A genera- tive flow for text-to-speech via monotonic alignment search,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.274609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.161679Z digest=sha256:ed2a842cf792afeaf72e9f80f0bf652fb68ab716b6b5e352ba24ab9235ba79f3

Observation 9b58995a-8d18-4580-8fb0-eef9d33e24f1 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.093806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.220293Z digest=sha256:5c1f3c7182cca3bc57b2233b722857fc4f54ff06cfa6e66e84451c76795f85a8

Observation 0e9bb990-4e1c-413b-9da5-0f0202d80a17 · outbound

This paper cites Grad-tts: A diffusion probabilistic model for text-to-speech,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Grad-tts: A diffusion probabilistic model for text-to-speech,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.897813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.285892Z digest=sha256:145d5074d94b8420ec1c258a9e733a2fda06341d7605b26ec7f54381c2039f22

Observation aa80d873-4dbe-4f85-a15b-fa7f44ee11b1 · outbound

This paper cites Msemotts: Multi-scale emotion transfer, prediction, and control for emotional speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Msemotts: Multi-scale emotion transfer, prediction, and control for emotional speech synthesis,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.642831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.800700Z digest=sha256:39aa4c40d187703f4ae85ba2c916ae853cf035dca17faf53e2e04932faa406fd

Observation dddd2a01-6871-4070-9e94-e000cc52b04b · outbound

This paper cites an unresolved cited work.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:00:39.558958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.410247Z digest=sha256:82bce282b561e8ba4ed49b9fcb8e75a306d4b61ed07ac7165d1b27cf6b42c706

Observation 284d4e11-006c-4139-af63-d30e9848ab6e · outbound

This paper cites Exploring transfer learning for low resource emotional tts,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Exploring transfer learning for low resource emotional tts,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.338794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.487306Z digest=sha256:f6ba3f6eb0b1c0ac1c092c75e0321b28e958fd572a97777215a7c95f3e4f4683

Observation ea7fac22-6097-4904-8b33-7671fa768928 · outbound

This paper cites Emospeech: guiding fastspeech2 to- wards emotional text to speech,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emospeech: guiding fastspeech2 to- wards emotional text to speech,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.186750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.544569Z digest=sha256:3f6209d6b120f827ecfa7b7a166a9f6dd7e612176134bd6be9ab044b8764f14c

Observation 6bc51587-e293-4e72-9edf-f9b106b4b5e8 · outbound

This paper cites Emodiff: Intensity con- trollable emotional text-to-speech with soft-label guidance,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emodiff: Intensity con- trollable emotional text-to-speech with soft-label guidance,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.018849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.650044Z digest=sha256:ef201dd835bd52c7132ff50a59677657d804751170f84e459644151c5928723f

Observation be948db1-5421-4a71-8889-1a531c49bb20 · outbound

This paper cites Controllable emotion trans- fer for end-to-end speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Controllable emotion trans- fer for end-to-end speech synthesis,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.840631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.743601Z digest=sha256:ba9300d4065f8e5790868e53566469daae3620a962f817a3a7f25648d549dd96

Observation f947fe3f-0bec-446f-8a11-b72a35a6b2bf · outbound

This paper cites Supervised and unsupervised approaches for controlling narrow lexical focus in sequence-to-sequence speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Supervised and unsupervised approaches for controlling narrow lexical focus in sequence-to-sequence speech synthesis,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.484831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.232683Z digest=sha256:2d2b7b46f17a58c12d5ea0a38abca6bec119e6c22fb43c5174eb2c918d575d25

Observation ef779f2c-c39b-4297-9dd7-dec3972d390c · outbound

This paper cites Daft- exprt: Cross-speaker prosody transfer on any text for expressive speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Daft- exprt: Cross-speaker prosody transfer on any text for expressive speech synthesis,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.455301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.890240Z digest=sha256:9f860bf143ffdcd25152ccca5c87e0e6989e1fe18273599426fb39203b7148ac

Observation be7b3e6e-a423-4945-8658-42504f261aa6 · outbound

This paper cites The acoustics of word stress in en- glish as a function of stress level and speaking style,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis The acoustics of word stress in en- glish as a function of stress level and speaking style,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.272357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:33.967567Z digest=sha256:78b89d9b26cef84bd88814fc4146ab6d6e33b9e9f7e88e2f248009bda4de3314

Observation c911473b-ea8e-4858-b756-d0855fd19cd0 · outbound

This paper cites The acoustics of lexical stress in italian as a func- tion of stress level and speaking style,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis The acoustics of lexical stress in italian as a func- tion of stress level and speaking style,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.036080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.027610Z digest=sha256:60d792b22078c9cbe082e81c0de96258171e681c915cb69cfb30e502d89218ec

Observation a2e2b0c8-7fce-49fa-8e04-6ea76cd34714 · outbound

This paper cites Lex- ical stress perception as a function of acoustic properties and the native language of the listener,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Lex- ical stress perception as a function of acoustic properties and the native language of the listener,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.842467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.084896Z digest=sha256:a61ca5b469419024617bf1292deab523108c6895ec59ee82bf2329bc7869d182

Observation 1fca9010-44b4-40c8-89b5-f1aab7da3eb5 · outbound

This paper cites Emphatic speech generation with conditioned input layer and bidirectional lstms for expressive speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emphatic speech generation with conditioned input layer and bidirectional lstms for expressive speech synthesis,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.669032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.167174Z digest=sha256:01729987c962e68df8f688f2d70805f4f23ab2de2da3cf75a1dcbd409c690e69

Observation 79660d22-0c18-4033-a8fb-aab94c464fe3 · outbound

This paper cites an unresolved cited work.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:00:41.173964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.662064Z digest=sha256:c9c21e4e1c8bbb1c3aa62b4e81f903dc08aa027b07f0bed636de8492205b3784

Observation 259c828e-80bb-444b-b2a7-34632b3bd53e · outbound

This paper cites Emphasis control for parallel neural tts,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emphasis control for parallel neural tts,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.315296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.332717Z digest=sha256:b59b95603872d33db53a77ed93519d975bfc18d4dade93bafc35ef8ae3c2faf6

Observation f819b3fa-7921-42ac-8baf-696418f5d8db · outbound

This paper cites Ee-tts: Emphatic expressive tts with linguistic information,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Ee-tts: Emphatic expressive tts with linguistic information,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.146568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.389567Z digest=sha256:eb309bd821d2435b04844be9186ccbcd3b3a3ebca93353aa8a8fd84f3585a7bb

Observation eefe6eab-eba2-49eb-b8e6-393ef89e99f4 · outbound

This paper cites A Survey of Large Language Models.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis A Survey of Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:34.450317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:34.450317Z digest=sha256:d4725b5b3f1c6dc6a9ca6d1147d4926bad9ea9811a3c00d75618170c57f2c32f

Observation b75f85c4-4b12-4115-8b86-c3b7c34fd809 · outbound

This paper cites Em- phAssess : a prosodic benchmark on assessing emphasis transfer in speech-to-speech models,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Em- phAssess : a prosodic benchmark on assessing emphasis transfer in speech-to-speech models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.941557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.527568Z digest=sha256:18a7f2c541201d11f37614934afa043200dec391639b7e324268a0fe6c6bb537

Observation 2d4b70b6-65c2-4e28-b4b3-948dc72dbeb6 · outbound

This paper cites Emotional voice con- version: Theory, databases and esd,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emotional voice con- version: Theory, databases and esd,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:34.632284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:34.632284Z digest=sha256:214549fdafe6f1d21420fcaad6851fd52ca7efa0343362ab18987fe36dcafc84

Observation 6e06e26b-7eab-48bf-924f-9882a5c7187b · outbound

This paper cites Hierarchical rep- resentation and estimation of prosody using continuous wavelet transform,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Hierarchical rep- resentation and estimation of prosody using continuous wavelet transform,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.727689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.713790Z digest=sha256:ab5e35cb6bbe19684c4022f3683f3b837560ec13499e1e196c8fe83976c45f9c

Observation 54819339-7cbf-4ad9-b4d7-a7ce54e7406f · outbound

This paper cites Adaspeech 4: Adaptive text to speech in zero-shot scenar- ios,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Adaspeech 4: Adaptive text to speech in zero-shot scenar- ios,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.504645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.803675Z digest=sha256:1cfaa5736b782a50103119e9ef0ff121656adc39d643b482fe8b2ff84a387443

Observation ea73529f-bfe4-47af-a1af-711f4ba6092e · outbound

This paper cites istftnet: Fast and lightweight mel-spectrogram vocoder incorporating inverse short-time fourier transform,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis istftnet: Fast and lightweight mel-spectrogram vocoder incorporating inverse short-time fourier transform,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.326584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.887690Z digest=sha256:1ffcb8c3ea324a7845e635ff8af0252dca98c2db49f79e8026c1f63dca256d20

Observation 2acf7453-1155-4e6c-b3e4-534d2479583e · outbound

This paper cites Adam: A method for stochastic optimization,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Adam: A method for stochastic optimization,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.152787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:34.970839Z digest=sha256:265d74b77abcd14ee3ca3b23f696bf7674c5b50cc2b832f5ff2bd073d7d65b7f

Observation 9df9ccfb-e94f-4307-b1ac-b3977703d489 · outbound

This paper cites CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:35.058089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:35.058089Z digest=sha256:52e57c1a9af80e237bf8e2d4dbf0763e45e6a5e0ebe8b9c496ce11b238a4f47d

Observation d419d07b-6e0d-42ae-aee4-43e7d537a330 · outbound

This paper cites GPT-4 Technical Report.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis GPT-4 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:35.146817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:35.146817Z digest=sha256:2f8aad0f90c9dee72b2cba1206fed914636edb221b21fd73e7c2ded8a737efd9

Observation 9a5cd7ae-cd45-4b15-a515-db92184589a8 · outbound

This paper cites emotion2vec: Self-supervised pre-training for speech emotion representation,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis emotion2vec: Self-supervised pre-training for speech emotion representation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.008196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:35.245582Z digest=sha256:6c710ab83d631103dc4aa987eb7a1c1c11b488c701ea0cea7d27fd2d22b62c70

Observation 3c4a8d21-99e2-4594-b225-9a385ed9ef18 · outbound

This paper cites Deep learning based assessment of syn- thetic speech naturalness,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Deep learning based assessment of syn- thetic speech naturalness,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:35.809221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:35.341553Z digest=sha256:64e2d24d8159078a1c322f3bfba5c26add2899bded8c0258a84cf33902dd9370

Pith citing papers

Observation a5be3642-8335-4aa2-bbcd-e0705651b4ea · inbound

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis cites this paper.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:00:35.612183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T17:00:32.452371Z digest=sha256:ae3d070b278150d6be37d9b3b32929eef93c9d272bdf8d19f2402e876b865e2c