Pith. sign in

Paper Citation Record · LEDGER

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

As of 9 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 3 inbound Pith citation observations for arXiv:1907.04448.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1907.04448 v2

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-25T00:09:35.661121Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:29.459675Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-25T00:10:07.071622Z

Reference resolution

36 of 36 outbound references displayed

  • verified exact5
  • verified fuzzy27
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ae3d3a5-6cbd-4435-b177-51c8f00bf4b9 · outbound

This paper cites prosody, by conditioning synthesis on la- tent representations [8–12] in addition to text.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning prosody, by conditioning synthesis on la- tent representations [8–12] in addition to text

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.660334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:15be54a031736f3681f7031fa11332c8e2ef3e1acac27f9ff402970f65b2d736

Observation 6fd71456-8b9d-4435-b97c-96a539929a35 · outbound

This paper cites Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T00:10:07.075022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:e973a6a488d5613f56823de5bbe2fee24fc34459e5a1049e99ae9508017cabe4

Observation 6cfd8f0d-51a0-4201-b5a4-346c0b1007e9 · outbound

This paper cites an unresolved cited work.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:10:07.642502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:80fb656c44bfb535ed6c8f83100d1cfecef08d598777c8cddcd2bea2d534a936

Observation fa48e1d4-4bc6-477f-925c-40da5a1e5ccf · outbound

This paper cites heavyaccented.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning heavyaccented

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.589630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:868658f027fc956e8eee182916ee332d8f591b99f7d0f631ad054981fa3ac87c

Observation 0cb291dc-c2f5-4781-906a-d237027c3a92 · outbound

This paper cites an unresolved cited work.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:10:07.648339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:a2b31eac740812260607e2bfcfd50e4bb3296adcc934daf789479bc3fdee9793

Observation 13dd437f-953e-4e43-9586-426de24c3f0f · outbound

This paper cites an unresolved cited work.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:10:07.680552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:197643c5bdabba51bc17e0d47d67e884dea3d2e7b064c9ed26142be1e9f2cc41

Observation dad61d25-0bb5-4b23-9366-265181562df0 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning WaveNet: A Generative Model for Raw Audio

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.041425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:0e3fc3c2430ca56fd2e648db34e698d54f7ad83b604cd8ae849c31da1ec37fc2

Observation da5c7525-49ce-40ba-b433-1f8767d18ee9 · outbound

This paper cites Tacotron: A fully end-to-end text-to-speech synthesis model.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Tacotron: A fully end-to-end text-to-speech synthesis model

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.576486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:293017e33fb6217007cacfc8eb881544898faad44dd3fdbcf579860b1ea27d7c

Observation 4a71613e-5913-4fa7-8311-5a843e6d3c49 · outbound

This paper cites DeepVoice2: Multi-speakerneuraltext- to-speech.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning DeepVoice2: Multi-speakerneuraltext- to-speech

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.572278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:a76996bcaf7fa52b76a94c1ea401dad2b9b51650d64a72d14c425c57969dc6db

Observation b6187871-b060-4c34-adce-e4132faea9e0 · outbound

This paper cites Neuralvoice cloning with a few samples.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Neuralvoice cloning with a few samples

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.667956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:7ef8a1aaee30000ab7212b109a5d67770a7d766d053ffc125866033ed2b8f164

Observation f73a218d-7ba3-46ec-ba2c-d62a60168311 · outbound

This paper cites Transfer learn- ing from speaker verification to multispeaker text-to-speech syn- thesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Transfer learn- ing from speaker verification to multispeaker text-to-speech syn- thesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.602137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:0676f90f61bd2169c777aee1239ae5bea20264cd406790a41f2d9ac2f1fe65e5

Observation 463ce1de-833a-4c65-bdd9-02d7ffd90e7f · outbound

This paper cites Fitting new speakers based on a short untranscribed sample.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Fitting new speakers based on a short untranscribed sample

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.672328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:e80e25d14eb61556b2ae7a915b8dc6605529fc1e58783be8802a76ab1fa99d91

Observation 50f1348c-36f9-4db1-9bd1-b67c37892f87 · outbound

This paper cites Sample Efficient Adaptive Text-to-Speech.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Sample Efficient Adaptive Text-to-Speech

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.062820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:decb6118fb2e54719c9ae29289f5cebcde90ccec71556ce9b50b0b9167095b94

Observation 1546fa51-115f-4a65-97e2-6a89c39779eb · outbound

This paper cites Style tokens: Unsupervised style modeling, control and transfer in end-to-end speechsynthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Style tokens: Unsupervised style modeling, control and transfer in end-to-end speechsynthesis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.605944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:eae5e90500d4541b8f3c4e8e4f79c6e1d259267de59a871e7cf2b31104be1db0

Observation 4e89d297-4da6-4fea-a261-56f9edce538e · outbound

This paper cites Towards end- to-endprosodytransferforexpressivespeechsynthesiswithTaco- tron.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Towards end- to-endprosodytransferforexpressivespeechsynthesiswithTaco- tron

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.618551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:8654abab1fde6b1ae9e8286949af9e34b39dbfe13e83e232da6f4799ca867e30

Observation 82dfebc3-3a6a-44a0-a101-eff418e3ecf1 · outbound

This paper cites Expressive speech synthesisviamodelingexpressionswithvariationalautoencoder.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Expressive speech synthesisviamodelingexpressionswithvariationalautoencoder

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.626764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:2990351571b8421a12ccf7351800be41359b079b4f2893902a360fef26e5616e

Observation 9ba06a2b-ac69-44c8-8be9-d88c3e4cc5f6 · outbound

This paper cites Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.055571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:6dc1e449f1c56a7c09968f10f2a51ddf113ee6e9310b432e6dfb4405f0d0ee1c

Observation 743c9309-9d1c-40f6-af1a-c75a2ed2f734 · outbound

This paper cites Hierarchical generative modeling for controllable speech synthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Hierarchical generative modeling for controllable speech synthesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.594414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:2741aa452b6c74079fed193a619d48d91ab3b3fc3a4d6f7eecfdba84a41ff0d1

Observation 60b12332-d185-49e3-8474-e24b70eafd40 · outbound

This paper cites Statistical parametric speech syn- thesis based on speaker and language factorization.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Statistical parametric speech syn- thesis based on speaker and language factorization

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.598354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:e478f1ed3728ff9e148c84ef6b66b1782a7125bb79a44611279fb29d92488e67

Observation bcdd3e0e-eea6-43b6-b75f-ad844e25f929 · outbound

This paper cites Multi-language multi-speaker acoustic model- ingforLSTM-RNNbasedstatisticalparametricspeechsynthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Multi-language multi-speaker acoustic model- ingforLSTM-RNNbasedstatisticalparametricspeechsynthesis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.609809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:2d77046a4e8ac38945092777b1d33d786d5fdd17cde8c037e9488b0aebe55fd9

Observation fa1e41ff-b3e3-4599-ba23-1f2a46f9ef41 · outbound

This paper cites A light-weight method of building an LSTM-RNN-based bilingual TTS system.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning A light-weight method of building an LSTM-RNN-based bilingual TTS system

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.614306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:b15ea59f364d4e56669ca6a13fb16deb83e923831da319ef36d79229d7317e77

Observation 0835da2a-2a5b-47f4-a6b5-5bf4cae7d335 · outbound

This paper cites Learning pronunciation from a foreign language in speech synthesis networks.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Learning pronunciation from a foreign language in speech synthesis networks

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-25T00:10:07.048463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:b95b67141e1c9c33315d14c6174fa3a5d1110f938026a9fb1597c4077f4a1f08

Observation 9cb4e778-c9bb-4a9f-8c20-c173c7e27d25 · outbound

This paper cites Unsupervisedpolyglottexttospeech.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Unsupervisedpolyglottexttospeech

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.676164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:fb903bf1e0b91d0526475bcc0d2f5e94f30ca23caf49a66d5711f5d698bc5174

Observation 343089da-d838-405c-abdc-153c89ec3e02 · outbound

This paper cites WORLD: a vocoder- based high-quality speech synthesis system for real-time applica- tions.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning WORLD: a vocoder- based high-quality speech synthesis system for real-time applica- tions

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.664180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:ca97caed5ad147367de194fb1937622206c7a32826d179bdfccc8beedf613157

Observation 21e1c771-0432-47df-b516-f1a73f85d01b · outbound

This paper cites Bytesareallyou need: End-to-end multilingual speech recognition and synthesis with bytes.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Bytesareallyou need: End-to-end multilingual speech recognition and synthesis with bytes

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.638343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:dae619b6a29e5de977a990cca87aff0f85c712016959374d6dc929f07d80f785

Observation 0c01e1a0-242a-4369-9fc8-35d5b94496fe · outbound

This paper cites Natural TTS synthesis by conditioning WaveNet on mel spectrogram predic- tions.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Natural TTS synthesis by conditioning WaveNet on mel spectrogram predic- tions

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.652174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:244a82fdffca4b90978cc4efe98ad56a2e3eb0efbfc39d1201733434198d4027

Observation 1ef0ec5c-fd0e-462e-b864-08100d5ee0b7 · outbound

This paper cites Auto-encodingvariationalBayes.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Auto-encodingvariationalBayes

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.656015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:0db9865fef804a2196d64947a5fdafc653037abc7a5eab2d19178a7e9637aabb

Observation ca15ffc2-a01b-4c46-9772-50af92048019 · outbound

This paper cites Efficient neural audio synthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Efficient neural audio synthesis

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.695580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:f3f38f8fcfc419d6f7276a1fd152eaf6a33d57c671fcfd144ded3e4e9148624a

Observation 586a05dc-7512-40bb-af36-0f5f68adfad3 · outbound

This paper cites Char2wav: End-to-endspeechsyn- thesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Char2wav: End-to-endspeechsyn- thesis

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.622605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:cd5fc4986ef53419d387267366fd22a7d3980026e08ae82c07c9aa9324ada175

Observation 79bd0c19-510b-4d92-b149-7d37f90d1262 · outbound

This paper cites Deep Voice 3: Scaling text-to-speech with convolutional sequence learning.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Deep Voice 3: Scaling text-to-speech with convolutional sequence learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.684223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:504d006fa382bd9ae4102951d8dd81b2c41528c48ee0ca0e6577ed86eff11935

Observation 6f3d1a08-bbbf-4049-b058-c4a0b3db7883 · outbound

This paper cites Representation Mixing for TTS Synthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Representation Mixing for TTS Synthesis

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.068668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:b4a7b0c7bb1c545c6a1473b35a92411e66c0164119645fc8494b96bf4e845054

Observation bc52e1f1-98cb-413c-8928-22b3002d98e4 · outbound

This paper cites Data-oriented methods for grapheme-to-phoneme conversion.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Data-oriented methods for grapheme-to-phoneme conversion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.631868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:210865555a817b52609430c8430eab2ba6251101eaafd8966eae6f5a1c3814b7

Observation 99409a6a-0220-44ba-8d48-ba0d1edc939b · outbound

This paper cites Domain- adversarial training of neural networks.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Domain- adversarial training of neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.687837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:75ce905c9dd1464b609773bbe7433226d76c12b838d1cb3b26352ef7e7b17c82

Observation d10ac97c-0db0-4a90-bd2b-2e4c23a7c895 · outbound

This paper cites Disentangling correlated speaker and noise for speech synthesis via data augmentation and adversarial factor- ization.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Disentangling correlated speaker and noise for speech synthesis via data augmentation and adversarial factor- ization

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.691566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:0c51a9f0ddca43627a0d4d211457cb3160914774f430d34f8b7d7cfba13e3ba2

Observation bd3acc62-3653-41ab-9dda-91fbb8c19d58 · outbound

This paper cites Cross-lingual speaker discrimination usingnaturalandsyntheticspeech.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Cross-lingual speaker discrimination usingnaturalandsyntheticspeech

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.584968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:84f3b595822149529cb4ef2067a1da672d2c230379938f621ebddb3ef1d33724

Observation 900e2af5-09e5-4d88-b18d-184bac19d2a1 · outbound

This paper cites Generalized end- to-end loss for speaker verification.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Generalized end- to-end loss for speaker verification

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.580729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:ce87ba35949a5b5457d8ce7ef62ddcea689e461ceaa89ea14c522749072120c7

Pith citing papers

Observation 6fd71456-8b9d-4435-b97c-96a539929a35 · inbound

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning cites this paper.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T00:10:07.075022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:e973a6a488d5613f56823de5bbe2fee24fc34459e5a1049e99ae9508017cabe4

Observation d54c8481-fba0-4c4e-bfaa-35dfdcb78861 · inbound

Optimizing Multilingual Text-To-Speech with Accents & Emotions cites this paper.

Optimizing Multilingual Text-To-Speech with Accents & Emotions Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:29.459675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:29.459675Z digest=sha256:254454f738a24329cb157790d70d80a60b023c7c98d5268b181392c90df3c970

Observation b5f15223-498d-4312-b7d4-6d7df3408c0e · inbound

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation cites this paper.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.408435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.408435Z digest=sha256:849866967507669b7c750d8641a763cbdfcaa1336a672163d418de357fbfd666