Pith. sign in

Paper Citation Record · LEDGER

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis

As of 20 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2507.12015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.12015 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:00:35.341553Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:00:32.452371Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:00:35.528453Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a5be3642-8335-4aa2-bbcd-e0705651b4ea · outbound

This paper cites EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:00:35.612183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.452371Z digest=sha256:e183a02caabccb86f0345dfeb6d782fda3e1995696d269598f5b60e69a426eb5

Observation f3fd56f6-044e-4b73-883b-f093917cbfa0 · outbound

This paper cites Overview The overall architecture of EME-TTS is shown in Figure 1a.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Overview The overall architecture of EME-TTS is shown in Figure 1a

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:41.294741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.546382Z digest=sha256:19b59f9b8fac54c29b8634d086d830a1c744b91d3fa83c30ad15652f6e600506

Observation ac11db75-a129-483a-a4e5-ffa6d14a7314 · outbound

This paper cites an unresolved cited work.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:00:41.067248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.720040Z digest=sha256:ef610597ebbbc2386fbb0e62f0aec1ce5eb0694bd1e5793c848228662e2b6b20

Observation e48b24f3-0364-4ccd-8c9a-fde12479c565 · outbound

This paper cites an unresolved cited work.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:00:40.949596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.761675Z digest=sha256:b0232ded1b8d6b653e9ae2aa36f1e1019774969d13dcb9a5925bb7ec961c8b5b

Observation e559f109-d637-4a3a-879f-2e350a30d97a · outbound

This paper cites LQN25F020001), and in part by the Key R&D Program of Zhejiang (2025C01104).

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis LQN25F020001), and in part by the Key R&D Program of Zhejiang (2025C01104)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.797255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.787020Z digest=sha256:72c5206c7bfb9d0855e74767c45f819e5dda2b411d677cbd0bb87a158629fb97

Observation 6a038c01-4ab0-470b-a154-038b0b632b2f · outbound

This paper cites Naturalspeech: End-to-end text-to- speech synthesis with human-level quality,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Naturalspeech: End-to-end text-to- speech synthesis with human-level quality,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.711282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.344834Z digest=sha256:1e970ce8300dd680c51d1c9c366300e5e6871905f63d762676e5b7a74b344c1b

Observation 6b0b10a4-0295-460e-83e8-27b08c43b8e7 · outbound

This paper cites Attention is all you need,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Attention is all you need,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.623965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.844343Z digest=sha256:48fcab6034de38cd292b989995aa9823b972acd21884a8803d3b5862c50c76ff

Observation 86a78097-fab6-45f8-8473-3cf0aa740f55 · outbound

This paper cites Fastspeech 2: Fast and high-quality end-to-end text to speech,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Fastspeech 2: Fast and high-quality end-to-end text to speech,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.424716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.970639Z digest=sha256:cc024e83aaca8f8e0c249a07c927a29a8acaa8f8b8322f12c97ecd3aa4af0013

Observation 1bb5d809-8fe0-4f8a-a5d7-4eb3638ff385 · outbound

This paper cites Glow-tts: A genera- tive flow for text-to-speech via monotonic alignment search,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Glow-tts: A genera- tive flow for text-to-speech via monotonic alignment search,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.274609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.161679Z digest=sha256:20ccc28184fa9de2589e6413aaf010554ce4e5d99aaa3a312d9af4ba0d09f72a

Observation 9b58995a-8d18-4580-8fb0-eef9d33e24f1 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:40.093806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.220293Z digest=sha256:ddce1802a1e112f3f52f0e5ccad0e830480d68a781abf4669118acf34de288cb

Observation 0e9bb990-4e1c-413b-9da5-0f0202d80a17 · outbound

This paper cites Grad-tts: A diffusion probabilistic model for text-to-speech,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Grad-tts: A diffusion probabilistic model for text-to-speech,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.897813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.285892Z digest=sha256:46385e58e284b17961e76f643dff73d70e957b5e68421482a4284b4d1c77e71d

Observation aa80d873-4dbe-4f85-a15b-fa7f44ee11b1 · outbound

This paper cites Msemotts: Multi-scale emotion transfer, prediction, and control for emotional speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Msemotts: Multi-scale emotion transfer, prediction, and control for emotional speech synthesis,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.642831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.800700Z digest=sha256:b5311a4fa47b303e1e71ec7a23067bce92e3db5d61a7d5d4bca6296ddb384eca

Observation dddd2a01-6871-4070-9e94-e000cc52b04b · outbound

This paper cites an unresolved cited work.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:00:39.558958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.410247Z digest=sha256:ab1090df274620bc040051c0742d59fa61aeeeebb5ee969c3958432dfcd09567

Observation 284d4e11-006c-4139-af63-d30e9848ab6e · outbound

This paper cites Exploring transfer learning for low resource emotional tts,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Exploring transfer learning for low resource emotional tts,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.338794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.487306Z digest=sha256:af269f2e642e897ae5d896c99b8339d8a483f9b6c2dd1db7e6b69039f60f5dc6

Observation ea7fac22-6097-4904-8b33-7671fa768928 · outbound

This paper cites Emospeech: guiding fastspeech2 to- wards emotional text to speech,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emospeech: guiding fastspeech2 to- wards emotional text to speech,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.186750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.544569Z digest=sha256:59c16fa0ca1b2d1b19f2c4f1b43109e57b1aced5e44597b7ba9061c27e5851ff

Observation 6bc51587-e293-4e72-9edf-f9b106b4b5e8 · outbound

This paper cites Emodiff: Intensity con- trollable emotional text-to-speech with soft-label guidance,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emodiff: Intensity con- trollable emotional text-to-speech with soft-label guidance,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:39.018849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.650044Z digest=sha256:ee025c75aa8f7da39fedbe3ac817e95dcf15ab451b4934edd78ac68574baead6

Observation be948db1-5421-4a71-8889-1a531c49bb20 · outbound

This paper cites Controllable emotion trans- fer for end-to-end speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Controllable emotion trans- fer for end-to-end speech synthesis,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.840631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.743601Z digest=sha256:31652ced112a7fa178e422f529dbda2b5a153c68bdc5bc838c3a9706bcc1ec9c

Observation f947fe3f-0bec-446f-8a11-b72a35a6b2bf · outbound

This paper cites Supervised and unsupervised approaches for controlling narrow lexical focus in sequence-to-sequence speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Supervised and unsupervised approaches for controlling narrow lexical focus in sequence-to-sequence speech synthesis,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.484831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.232683Z digest=sha256:eacd0c6f1d25bd3d284aa7a8e8b412dde97ca84528038c576bca48e29f03714c

Observation ef779f2c-c39b-4297-9dd7-dec3972d390c · outbound

This paper cites Daft- exprt: Cross-speaker prosody transfer on any text for expressive speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Daft- exprt: Cross-speaker prosody transfer on any text for expressive speech synthesis,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.455301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.890240Z digest=sha256:52a09e57413c071fd541cc5a89d0a32ff11b6c10093ec8a10a9246eecfe29e79

Observation be7b3e6e-a423-4945-8658-42504f261aa6 · outbound

This paper cites The acoustics of word stress in en- glish as a function of stress level and speaking style,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis The acoustics of word stress in en- glish as a function of stress level and speaking style,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.272357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:33.967567Z digest=sha256:02ef506e1edec378a94d4f48d239ac2ebbe2d269a46416aff246b7ef8ab31e8a

Observation c911473b-ea8e-4858-b756-d0855fd19cd0 · outbound

This paper cites The acoustics of lexical stress in italian as a func- tion of stress level and speaking style,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis The acoustics of lexical stress in italian as a func- tion of stress level and speaking style,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:38.036080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.027610Z digest=sha256:ff9f51a95f86553e2332b65122badcec2c7e2acadfce12fad30c7454b0061132

Observation a2e2b0c8-7fce-49fa-8e04-6ea76cd34714 · outbound

This paper cites Lex- ical stress perception as a function of acoustic properties and the native language of the listener,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Lex- ical stress perception as a function of acoustic properties and the native language of the listener,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.842467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.084896Z digest=sha256:8252fe67f14aae327cc8bbdddf51bd974b8bbf5ac770c1ddc88a00e0d57b4849

Observation 1fca9010-44b4-40c8-89b5-f1aab7da3eb5 · outbound

This paper cites Emphatic speech generation with conditioned input layer and bidirectional lstms for expressive speech synthesis,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emphatic speech generation with conditioned input layer and bidirectional lstms for expressive speech synthesis,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.669032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.167174Z digest=sha256:c48dd2b4bcfb0711a71e4d6051e5ebe3b63340a1f3ae93886c0353cabd73b036

Observation 79660d22-0c18-4033-a8fb-aab94c464fe3 · outbound

This paper cites an unresolved cited work.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:00:41.173964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.662064Z digest=sha256:45c53619f825c8f07d655a3b33e0746557a58962d1915b12456ab475308f6b1e

Observation 259c828e-80bb-444b-b2a7-34632b3bd53e · outbound

This paper cites Emphasis control for parallel neural tts,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emphasis control for parallel neural tts,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.315296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.332717Z digest=sha256:822c78338fe147cd17ba0b3d06f08137a740f5c302f7b2dfbe17ea74a5ecdfe0

Observation f819b3fa-7921-42ac-8baf-696418f5d8db · outbound

This paper cites Ee-tts: Emphatic expressive tts with linguistic information,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Ee-tts: Emphatic expressive tts with linguistic information,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:37.146568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.389567Z digest=sha256:4b1a6029546ed75b6e7c30cebcbf5fb82e0137c67760f9c2d61b551974324c9d

Observation eefe6eab-eba2-49eb-b8e6-393ef89e99f4 · outbound

This paper cites A Survey of Large Language Models.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis A Survey of Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:34.450317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:34.450317Z digest=sha256:d4725b5b3f1c6dc6a9ca6d1147d4926bad9ea9811a3c00d75618170c57f2c32f

Observation b75f85c4-4b12-4115-8b86-c3b7c34fd809 · outbound

This paper cites Em- phAssess : a prosodic benchmark on assessing emphasis transfer in speech-to-speech models,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Em- phAssess : a prosodic benchmark on assessing emphasis transfer in speech-to-speech models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.941557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.527568Z digest=sha256:17f53afec20bc1f04a858fefee5b2a55bdbeb3b2d93fbee2d512c7d1211a8124

Observation 2d4b70b6-65c2-4e28-b4b3-948dc72dbeb6 · outbound

This paper cites Emotional voice con- version: Theory, databases and esd,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Emotional voice con- version: Theory, databases and esd,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:34.632284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:34.632284Z digest=sha256:214549fdafe6f1d21420fcaad6851fd52ca7efa0343362ab18987fe36dcafc84

Observation 6e06e26b-7eab-48bf-924f-9882a5c7187b · outbound

This paper cites Hierarchical rep- resentation and estimation of prosody using continuous wavelet transform,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Hierarchical rep- resentation and estimation of prosody using continuous wavelet transform,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.727689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.713790Z digest=sha256:2b0a4476e06d76c4d4d76bf502c67282ee725a5f8b55ffb5f14e99dbff5dd0c4

Observation 54819339-7cbf-4ad9-b4d7-a7ce54e7406f · outbound

This paper cites Adaspeech 4: Adaptive text to speech in zero-shot scenar- ios,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Adaspeech 4: Adaptive text to speech in zero-shot scenar- ios,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.504645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.803675Z digest=sha256:9754fcf10b0826d77af41711831e0b77d119a542551eb345379b864d30ad1bcc

Observation ea73529f-bfe4-47af-a1af-711f4ba6092e · outbound

This paper cites istftnet: Fast and lightweight mel-spectrogram vocoder incorporating inverse short-time fourier transform,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis istftnet: Fast and lightweight mel-spectrogram vocoder incorporating inverse short-time fourier transform,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.326584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.887690Z digest=sha256:8013ee3ae1423a34bbf4981e0115d8e46367076ad70641fb37a2b23c9a2b31d7

Observation 2acf7453-1155-4e6c-b3e4-534d2479583e · outbound

This paper cites Adam: A method for stochastic optimization,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Adam: A method for stochastic optimization,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.152787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:34.970839Z digest=sha256:7f7dfe2a205fd589242d17d633d7be9d0ba5c0ee7fed5ff4ff8eb824d604db7b

Observation 9df9ccfb-e94f-4307-b1ac-b3977703d489 · outbound

This paper cites CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:35.058089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:35.058089Z digest=sha256:52e57c1a9af80e237bf8e2d4dbf0763e45e6a5e0ebe8b9c496ce11b238a4f47d

Observation d419d07b-6e0d-42ae-aee4-43e7d537a330 · outbound

This paper cites GPT-4 Technical Report.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis GPT-4 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:35.146817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:35.146817Z digest=sha256:2f8aad0f90c9dee72b2cba1206fed914636edb221b21fd73e7c2ded8a737efd9

Observation 9a5cd7ae-cd45-4b15-a515-db92184589a8 · outbound

This paper cites emotion2vec: Self-supervised pre-training for speech emotion representation,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis emotion2vec: Self-supervised pre-training for speech emotion representation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:36.008196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:35.245582Z digest=sha256:49a0fcb72043804de09d57359bc954d3d7e21844e4fc618766a6fcc9b8a2eb4d

Observation 3c4a8d21-99e2-4594-b225-9a385ed9ef18 · outbound

This paper cites Deep learning based assessment of syn- thetic speech naturalness,.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis Deep learning based assessment of syn- thetic speech naturalness,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:00:35.809221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:35.341553Z digest=sha256:fd117349076dca258fdb924911d8b0ac3cc377b1c4e55e6f233fd2f653865ccd

Pith citing papers

Observation a5be3642-8335-4aa2-bbcd-e0705651b4ea · inbound

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis cites this paper.

EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:00:35.612183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:00:32.452371Z digest=sha256:e183a02caabccb86f0345dfeb6d782fda3e1995696d269598f5b60e69a426eb5