Pith. sign in

Paper Citation Record · LEDGER

EmoSpeech: Guiding FastSpeech2 Towards Emotional Text to Speech

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2307.00024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.00024 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:25:11.559191Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T21:19:03.234818Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 58dd6b1a-6de1-4649-8ea6-d008f05c8974 · inbound

A Review of Human Emotion Synthesis Based on Generative Technology cites this paper.

A Review of Human Emotion Synthesis Based on Generative Technology EmoSpeech: Guiding FastSpeech2 Towards Emotional Text to Speech

Reference 177

Resolution
unresolved
no resolver link, observed 2026-08-11T19:10:14.102836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:10:14.102836Z digest=sha256:f55a63d2bac80c93807bb563dbaf5494b556321915579185f393aa8518bfb8bf

Observation 24bf4161-6c7d-444c-93c0-c2f5c17b3be1 · inbound

MultiActor-Audiobook: Zero-Shot Audiobook Generation with Faces and Voices of Multiple Speakers cites this paper.

MultiActor-Audiobook: Zero-Shot Audiobook Generation with Faces and Voices of Multiple Speakers EmoSpeech: Guiding FastSpeech2 Towards Emotional Text to Speech

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:11.559191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:25:11.559191Z digest=sha256:6a01d943e854404004b3c9a308c8d567c40c79cd96b5c2f1edd6f76f25a39607

Observation a6328654-ac4e-4765-a3e6-51717d90e619 · inbound

Marco-Voice Technical Report cites this paper.

Marco-Voice Technical Report EmoSpeech: Guiding FastSpeech2 Towards Emotional Text to Speech

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T05:17:05.775601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:17:05.775601Z digest=sha256:45d89fc31ba0f07be307fdafc221d91e90d685a1d749b005b3a9e0d33e24b242

Observation 1183f891-e637-4e9c-bc8d-8eb45fdb4e99 · inbound

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech cites this paper.

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech EmoSpeech: Guiding FastSpeech2 Towards Emotional Text to Speech

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:19:03.236339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-20T21:14:58.814362Z digest=sha256:8277651878425c27f18d0e34e03cbc7e35e7947148cfd49cf7bcacccf1722c71

Observation 5677a6ab-52a9-4e6b-9959-4381e03d0394 · inbound

Beyond One-Size-Fits-All: Personalized and Culturally Adaptive Emotional TTS via Interactive Optimization of Individual Emotion Perception Spaces cites this paper.

Beyond One-Size-Fits-All: Personalized and Culturally Adaptive Emotional TTS via Interactive Optimization of Individual Emotion Perception Spaces EmoSpeech: Guiding FastSpeech2 Towards Emotional Text to Speech

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:31.803047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:31.803047Z digest=sha256:d6fce727ee0c71cd5b92b53eedecd87e4d6b9410fb2f51585d67d1f3715667b2