Pith. sign in

Paper Citation Record · LEDGER

EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2504.12867.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.12867 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:29.525601Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T02:49:24.567639Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 078c576c-a685-4c0e-aee4-d2d6c75b4979 · inbound

Optimizing Multilingual Text-To-Speech with Accents & Emotions cites this paper.

Optimizing Multilingual Text-To-Speech with Accents & Emotions EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:29.525601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:29.525601Z digest=sha256:fef89ca4fb8bf2e4929d442de0b6f1de7d90ce7c4799bd158a18fe6b5232a96e

Observation 69226588-46d7-464f-826b-30bc3caa6198 · inbound

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation cites this paper.

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:58.957288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:47:58.957288Z digest=sha256:4ab6aef1ed9f475c9f6855aa9aa58182a403c7c4e866df6b6ebcbf78753390e1

Observation e8030ec9-5c4e-4c8b-b9de-b51f1d09328a · inbound

MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts cites this paper.

MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T20:04:48.933028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:04:48.933028Z digest=sha256:f8bde3c828a8b305edd3553b80bce45764950ec308591b0ca471cd011ba3ff7f

Observation 8e58e325-e656-4e23-9637-e05c7bfe76a8 · inbound

Semantic-Aware Ship Detection with Vision-Language Integration cites this paper.

Semantic-Aware Ship Detection with Vision-Language Integration EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T17:41:58.565216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:41:58.565216Z digest=sha256:c0ef59aa824b10e3b0d1d71d06c85330e7e69132fbff3dd2591aad3fa32acf5d

Observation dd4eee83-4785-4c3c-8e77-e5edad12565d · inbound

Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning cites this paper.

Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:01:01.853953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:50:59.950646Z digest=sha256:9d42e1ede060376ca13b1a4e60ea3805338ea9a25607f02805ab2f9d71f2b6e5

Observation 4d172644-6192-4f82-b2db-ff95e80a089b · inbound

UniVocal: Unified Speech-Singing Code-Switching Synthesis cites this paper.

UniVocal: Unified Speech-Singing Code-Switching Synthesis EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T00:46:24.542484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T13:17:13.510587Z digest=sha256:bba5b6d4e15e52b826ab0f4f32c34b0bebdb6f9496ef653785d65cb36b204fe1

Observation ea96c1bd-869f-4af7-8758-c490e8ab2cf9 · inbound

FineCombo-TTS: Collaborative and Precise Controllable Speech Synthesis Using Text Descriptions and Reference Speech cites this paper.

FineCombo-TTS: Collaborative and Precise Controllable Speech Synthesis Using Text Descriptions and Reference Speech EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:49:24.569112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T19:09:33.605232Z digest=sha256:9d0aa46d5d1203a9529828364728157b7e4bb67ba3e064dd9e21d3c3ddaa611d

Observation 7ffc3279-1555-42dc-b45b-541efe36508d · inbound

EmoInstruct-TTS: Dual-Path Instruction-Guided Emotional Speech Synthesis cites this paper.

EmoInstruct-TTS: Dual-Path Instruction-Guided Emotional Speech Synthesis EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:47:30.568695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T17:01:13.972071Z digest=sha256:454edf25aeeaadd87ee7ceedbed92fc918ebc2b66534465df91072ded6538741