Pith. sign in

Paper Citation Record · LEDGER

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

As of 10 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 1 inbound Pith citation observation for arXiv:2507.02176.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02176 v1

Coverage vector

measured 69 of 69 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:40:16.437239Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:40:10.960719Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T20:40:17.636566Z

Reference resolution

69 of 69 outbound references displayed

  • verified exact4
  • verified fuzzy56
  • unresolved8
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fbd3ef56-0b7a-4d69-8241-d50fb04b296f · outbound

This paper cites Virtual chara cters are expected to possess unique identities that remain consi stent across time.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Virtual chara cters are expected to possess unique identities that remain consi stent across time

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:29.587245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:10.892442Z digest=sha256:dd0444bcd90b92f9abd296e12167421548753e1813869de9e788584e4c3c9e7f

Observation 517abf04-5934-49ef-95b2-e23ab2515c83 · outbound

This paper cites Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:17.833571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:10.960719Z digest=sha256:38d6c10e65aeeef8e404bc454946d097f97d795c4851f0e21dad1eba3db7d0ed

Observation f1f5725e-124c-47c6-963b-359f243ac434 · outbound

This paper cites We explore which mark- ers are represented in some widely used ASV embeddings, and measure the effect of confounding factors.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We explore which mark- ers are represented in some widely used ASV embeddings, and measure the effect of confounding factors

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.996348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.033440Z digest=sha256:6535243d5ba00862a7f71e493e4f0cc14953e6e41041f0e59b64d482f6553a7a

Observation ce4bbb54-3138-4eec-b835-4404f612eb6d · outbound

This paper cites This reflects the lower bound of our metric, where we expect the smallest distances.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis This reflects the lower bound of our metric, where we expect the smallest distances

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.311331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.119053Z digest=sha256:9f09f5af8ec299a91d2d9cfd333a555ee93bc215eee792eb768031ecd5300640

Observation 8bbfef3e-0a69-429e-9162-8340ecfa3475 · outbound

This paper cites This setting tests wh ether our metric can distinguish between speakers undistinguish - able with speech rate.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis This setting tests wh ether our metric can distinguish between speakers undistinguish - able with speech rate

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.102856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.225540Z digest=sha256:c5270a442716f8096ebe0768ccd0fed9a1c01e495c92c467adc783bd20567a16

Observation 89f98804-a7b9-49a1-99e0-d0bacc9f6e2c · outbound

This paper cites The results in the top section of Table 3 show that rhythm distances are significantly larger between different speak ers, even those with similar speech rates.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The results in the top section of Table 3 show that rhythm distances are significantly larger between different speak ers, even those with similar speech rates

Reference 6

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:40:27.992700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.368279Z digest=sha256:dc16fc3589fafdd0b70b070f27972e0c93f257bf7bd4ae28597b15dc8cd8bad0

Observation 0ec6aae4-9475-4471-a80b-ff699a3f4477 · outbound

This paper cites We showed that ASV embeddings mainly encode speech identity markers relating to anatomy (e.g.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We showed that ASV embeddings mainly encode speech identity markers relating to anatomy (e.g

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.880864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.464887Z digest=sha256:a03f9eb63580dc62420eec591a637b249e28b6a91f85830c7f893b60e681a251

Observation 5d5d9667-5627-4180-8746-56411436c248 · outbound

This paper cites An image is worth one word: Personalizing t ext-to- image generation using textual inversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis An image is worth one word: Personalizing t ext-to- image generation using textual inversion,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.756865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.537350Z digest=sha256:4bc744636d7f8fe0ed11eefba0f9f4583e60e557ea786805f5d34aff4c841c64

Observation 932e5017-ad30-4164-a120-a6d46a36daa7 · outbound

This paper cites Dreambooth: Fine tuning text-to-image d iffusion models for subject-driven generation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Dreambooth: Fine tuning text-to-image d iffusion models for subject-driven generation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.588974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.595403Z digest=sha256:95920ef2fd39fe94d475722c6868f928e87ac5602348f6649a0482ef4b118706

Observation 0c62ff45-0144-4f18-aa42-7e8b96b9b4f1 · outbound

This paper cites ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:11.679021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:11.679021Z digest=sha256:da6991931092f7476d1156ddb2d7baeae64380485f894713ea456ee364390dbd

Observation c89f5a9f-9e70-4ff9-a821-38f5722804a4 · outbound

This paper cites CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:11.744848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:11.744848Z digest=sha256:abc30cfd59331ced0c9848e2f34f7acc71b5d46176901d5205e5316b78f7a5e8

Observation 5d17b08b-2cef-42f0-b20c-065a1b807fe3 · outbound

This paper cites Generat ing video game scripts with style,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generat ing video game scripts with style,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.458551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.852395Z digest=sha256:793bd4763ece28bd29e5b501043e845fa4ac6a582019cde54a2a360b5ff3addf

Observation 988925a3-6791-4cc9-b3ea-6d3176cae942 · outbound

This paper cites Meet your favorite character: Open-domai n chat- bot mimicking fictional characters with only a few utterance s,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Meet your favorite character: Open-domai n chat- bot mimicking fictional characters with only a few utterance s,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.351960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.916503Z digest=sha256:4c2dedb91d9bc32478a1e2ab0bb6452cb33de66b744a2e542956cf42b776127b

Observation ba599d21-aa4a-44e9-9de5-a9c72b8282d5 · outbound

This paper cites Information conveyed by vow- els.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Information conveyed by vow- els

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.170164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:11.991064Z digest=sha256:9380b2370bec4cf5d40c67b984b752124ffb5586d4b96b843e11073ad6152096

Observation d6d6a64a-a140-4335-95bd-c7a4a8d02172 · outbound

This paper cites The perception of personal iden tity in speech: Evidence from the perception of twins’ speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The perception of personal iden tity in speech: Evidence from the perception of twins’ speech,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.987242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.078389Z digest=sha256:d5f75e2710c4568e464510f1a3e7e14f3a53ac0877d2f7177007f7981d88c4ff

Observation e7c45ab1-d99e-4c5a-812e-3435e2d3b198 · outbound

This paper cites X-V ectors: Robust DNN embeddings for speaker recognition,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis X-V ectors: Robust DNN embeddings for speaker recognition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.806842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.158681Z digest=sha256:a396788589204b055cad1801a13981d2c08f9211df6d7684fdfe2f9b833bac6d

Observation f7af280c-6e79-496f-a69c-c5d8babb5c99 · outbound

This paper cites ECAPA - TDNN: emphasized channel attention, propagation and aggre ga- tion in TDNN based speaker verification,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ECAPA - TDNN: emphasized channel attention, propagation and aggre ga- tion in TDNN based speaker verification,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.629537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.245018Z digest=sha256:3dbf616d4a713937fb56bfca5bc6ba531787245f462962a8d43cd92adba5f1f3

Observation cc153765-41c9-49ce-b995-37e9c0f1de4d · outbound

This paper cites State-of-the-art speaker recogni tion with neural network embeddings in NIST SRE18 and speakers in the wild evaluations,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis State-of-the-art speaker recogni tion with neural network embeddings in NIST SRE18 and speakers in the wild evaluations,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.482722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.319383Z digest=sha256:2f927ba4d560ab0da1aa03c712bbdfa9194224482f2830569158c68f7f3eeb55

Observation cb4977ac-8be1-44ea-ac18-7e54da19fefd · outbound

This paper cites Generalized end-to-end loss for speaker verifica- tion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generalized end-to-end loss for speaker verifica- tion,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.302718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.410015Z digest=sha256:26d4dcb5ee1df8012cea315fe6a0b7018a995ba84ab0edb3e497114852e951b8

Observation b5ee12e4-44fe-4f82-9642-408fa023c43c · outbound

This paper cites We need variations in speech synthe- sis: Sub-center modelling for speaker embeddings,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We need variations in speech synthe- sis: Sub-center modelling for speaker embeddings,

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-06T20:40:17.374269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.504831Z digest=sha256:8cb565ecfcbf77a9efbaa6a630f5cee4aa067f15366ef89f8c0cc89677c9d11c

Observation 7f5eebdb-5db5-4fd2-87f8-7537d1edc768 · outbound

This paper cites Predictions of subjective ratings and spoofing as- sessments of V oice Conversion Challenge 2020 submissions,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Predictions of subjective ratings and spoofing as- sessments of V oice Conversion Challenge 2020 submissions,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.121033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.603592Z digest=sha256:c50696faa0d5ae75b3fd4e7ca5b6f9a14572f8cbb75af9d4d95c095c5a901cbe

Observation b0d89eba-6eef-4933-97ed-e2d568c16f86 · outbound

This paper cites Transfer learning from speaker verificat ion to multi- speaker text-to-speech synthesis,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Transfer learning from speaker verificat ion to multi- speaker text-to-speech synthesis,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.902006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.679830Z digest=sha256:f9c3549366f824293fb2fb278119d58b00114e9540029a354236c80fd03a9ed0

Observation d59ce44c-d9f4-412b-bedd-6b5094c18c70 · outbound

This paper cites Zero-shot multi-speaker text-to-sp eech with state-of-the-art neural speaker embeddings,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Zero-shot multi-speaker text-to-sp eech with state-of-the-art neural speaker embeddings,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.728737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.740868Z digest=sha256:f8ab0cc26e12dcf2393a547151f5d3cf13187a30f8b65776ca649860e6952dfd

Observation 38425e51-2504-4a84-874e-c042e926c1d6 · outbound

This paper cites Y ourTTS: towards zero-shot multi- speaker tts and zero-shot voice conversion for everyone,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Y ourTTS: towards zero-shot multi- speaker tts and zero-shot voice conversion for everyone,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.600183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.837534Z digest=sha256:407ef35aff7912f8740e9489a3aa199f54550c6406260aa024d941b85da1c802

Observation 1cb6588a-a323-4f5c-b318-afc26a02eae4 · outbound

This paper cites Investigating on incorporating pret rained and learnable speaker representations for multi-speaker mult i-style text-to-speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Investigating on incorporating pret rained and learnable speaker representations for multi-speaker mult i-style text-to-speech,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.420748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:12.936260Z digest=sha256:e12a3ffd0649edab6e9688766a41f78bf406b9c8b786dd63c6b8984aafb3f726

Observation e3bf7555-52cb-4bf9-bd88-e292621223c0 · outbound

This paper cites Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.032052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.032052Z digest=sha256:79c99ea36aedcffc7d0b958ce048a62ae18631effce5a8c985b794bd3564fcec

Observation 01a602c7-7a3a-4992-bcca-2d20ea94f10d · outbound

This paper cites The Multi-Speaker Multi-Style V oice Clo ning Chal- lenge 2021,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Multi-Speaker Multi-Style V oice Clo ning Chal- lenge 2021,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.237363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.128895Z digest=sha256:12615b8249b94032789e0c23190e3be03c55e8b3e1b00da4980ed1632332486e

Observation 4857260a-648a-4d2c-80a3-34fd08fb25de · outbound

This paper cites VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.209463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.209463Z digest=sha256:64fe6fcfc69de4270639008d8813c7e3314857b1d0e83e6144aa2cbafaebe216

Observation fc46cf74-cdd9-420c-b7f8-d04b70d048b7 · outbound

This paper cites Evaluating text-to-speech synthesis from a large discrete token-based speech language model,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Evaluating text-to-speech synthesis from a large discrete token-based speech language model,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.020012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.330030Z digest=sha256:694ed362ee7e65563ebfb680e32abee1fb9d4dd3865b961b786d500ca19d5a23

Observation f6e68c49-0f0f-4721-b3a1-898330ee7271 · outbound

This paper cites A comparison of discrete and soft speech units for improved voice conversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis A comparison of discrete and soft speech units for improved voice conversion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.420536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.420536Z digest=sha256:600f21661c90d787a7f311e5efe41c72a24607ef3246b0d5adad0996c276b2ba

Observation 9fdce733-cd03-4761-b18d-ef7f05f91428 · outbound

This paper cites V oicebox: Text-guided multilingual univ ersal speech generation at scale,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V oicebox: Text-guided multilingual univ ersal speech generation at scale,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.864555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.484801Z digest=sha256:cbe84b884790a4e918f4754b7f298a9315696d19e6ea5e35f8b07fa4f714cfc4

Observation 67b8fbcc-5f6d-447d-bf18-4f153dc8ce3d · outbound

This paper cites Neural codec language models are zero-s hot text to speech synthesizers,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Neural codec language models are zero-s hot text to speech synthesizers,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.567138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.544795Z digest=sha256:e29524badb921beee9e8a8805c3897d998ba9d137a9648778a7f051d08a4a105

Observation fc5d7913-3f76-4168-90e3-16b024875a87 · outbound

This paper cites The Singing V oice Conversion Chall enge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Singing V oice Conversion Chall enge 2023,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.362172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.610042Z digest=sha256:6e4b523e40606987a06e3202d953939d6f21659423cf3cea6a4a187ea9a9bad3

Observation f3ef1777-0b8f-4ab0-9ddc-4580ee3c651c · outbound

This paper cites Generative Data Augmentation Challe nge: Zero- shot speech synthesis for personalized speech enhancement ,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generative Data Augmentation Challe nge: Zero- shot speech synthesis for personalized speech enhancement ,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.160692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.678518Z digest=sha256:99f64e534783131506c0ce014062883d56037dd52a34e002d8a68e0cb0e8f907

Observation 654da8cc-1fdc-4b24-b46c-06147b004493 · outbound

This paper cites Acoustic properties of voice timbre t ypes and their influence on voice classification,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Acoustic properties of voice timbre t ypes and their influence on voice classification,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.920263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.758589Z digest=sha256:ded23aa3680d4f82cbfed1ff4630ec2ec4360cee8a48405663135e48dfbb8976

Observation 5c79d98a-4962-4f10-802e-f487f2340f69 · outbound

This paper cites Speech production patterns in producing l inguis- tic contrasts are partly determined by individual differen ces in anatomy,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Speech production patterns in producing l inguis- tic contrasts are partly determined by individual differen ces in anatomy,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.680062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.823753Z digest=sha256:7c3add1023f1faeaf788da2bad4ba70bdb5f1c701f302bc62fde68d757afb558

Observation bc66290e-98e4-4bd5-981a-3190fe79369b · outbound

This paper cites an unresolved cited work.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:40:23.460801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.871818Z digest=sha256:0d0e9f7501580d9ad25e6f61ef1fa3fb43a9df92d19b916ea95f2bb3a6c8d26e

Observation cae77f0c-21b8-4b0b-8ad9-32e71334c683 · outbound

This paper cites A study of rhythm in London: Is syllable-timing a feature of Multicultural London English ?.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis A study of rhythm in London: Is syllable-timing a feature of Multicultural London English ?

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.289528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.928173Z digest=sha256:de0cbe0b593803a79635375055f4ac55774c860820915977179d0fadad9b6598

Observation af5d4057-27d1-48f6-961b-011f8958c4ba · outbound

This paper cites The measurement of rhythm: a comparison of Sin- gapore and British English,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The measurement of rhythm: a comparison of Sin- gapore and British English,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.067362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:13.978443Z digest=sha256:1bc703f68096a9d2a62c8c61d017e3cc4eb5d741f4bf8ce991fa30b6426d49eb

Observation 10783b19-ca12-4bc5-8432-b0df7d8a26ba · outbound

This paper cites Sociophonetics of phonotactic pheno mena in french,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Sociophonetics of phonotactic pheno mena in french,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.807797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.042470Z digest=sha256:ef84021efe3e3fa55a0e17b53ddc8136a6e6c8ec2f505a2fed43b49f408c783f

Observation 4d90a904-75f6-4cf3-8437-f4ba2f576871 · outbound

This paper cites Individual di fferences in vowel production,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Individual di fferences in vowel production,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.620821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.108230Z digest=sha256:6d988a5cd7c5d9e760dc7db2235b6e3e0ca2e3a0185c5e7812dea089d5771860

Observation 13627fac-ceff-4ca7-9de0-37f89d1c47df · outbound

This paper cites Individual differences in speech pro- duction: V oice-onset-time,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Individual differences in speech pro- duction: V oice-onset-time,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.428908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.167152Z digest=sha256:85bf1df60148b8dc45e072e8f5553ec886ffb93b9f724758fe6cc2e2a08b02a8

Observation 6c2ed39d-60f7-4358-bb49-930065f56271 · outbound

This paper cites The social life of phonetic s and phonology,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The social life of phonetic s and phonology,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.226893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.225970Z digest=sha256:dfd2dbcb0d32d497022dcc442396b51e38803845378b4c783210d4780b665f8a

Observation a0a78c3e-a471-4b32-9051-7867fd9eefd9 · outbound

This paper cites Prosodic rhythm and Afric an American english,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Prosodic rhythm and Afric an American english,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.055134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.273503Z digest=sha256:a5785063a47e30e5e98a00271002f43b7320caeaa20d47dfec85b4fd0369227a

Observation dd9e3107-7a7a-413a-82ef-74ec4916a3e9 · outbound

This paper cites La liaison sans enchaˆ ınement,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis La liaison sans enchaˆ ınement,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.865653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.325252Z digest=sha256:5d072c040ba0000df166f9cef25c7901c8dee938712c35a718e70939052a6a5b

Observation f7e19809-8f4a-47d6-b320-721f6b2d4a4b · outbound

This paper cites Stuart-Smith, E.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Stuart-Smith, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.707914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.381549Z digest=sha256:4ed8ab937aeeb90d2a63ea31cc807f1fe2479ce1e25db6bf3272fd039da058f8

Observation f2fee590-4f65-4a33-85bb-bb5f1f5ae99f · outbound

This paper cites The Blizzard Challenge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Blizzard Challenge 2023,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.492072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.437994Z digest=sha256:ff0ec6eafd0150fde91edc6098ed34b2b4cef93687fab1e3a12a96be380cb7aa

Observation 7c04876d-3c13-4c25-8171-85a9405c39fb · outbound

This paper cites Refining the evaluation of speech synthesis: A summ ary of the blizzard challenge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Refining the evaluation of speech synthesis: A summ ary of the blizzard challenge 2023,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.310864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.486042Z digest=sha256:464bda8e7530e4c18cb9431254294e5d8a02601b6c2003fb24ba40e146ff3270

Observation 457a1f57-16e9-4294-83bd-5d1340707f4c · outbound

This paper cites Good practices for evaluation of synthesized speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Good practices for evaluation of synthesized speech,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:14.536826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:14.536826Z digest=sha256:834b7d3ae1d4180f68295bb175b0223dc4311cf051a7be67eef2aed67ed34ae0

Observation a8594ffb-9d77-4924-a681-932f2c5a1c65 · outbound

This paper cites Stuck in the MOS pit: A critical anal ysis of MOS test methodology in TTS evaluation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Stuck in the MOS pit: A critical anal ysis of MOS test methodology in TTS evaluation,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.155580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.599962Z digest=sha256:a389674811da3be50fa985e5c9147baa971a8b84c39095e80717b4bd4df291b9

Observation bf0a2f81-a772-4739-9fd1-56a402a46427 · outbound

This paper cites Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:16.930605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.655226Z digest=sha256:61efe50ae1e145bce1521fce4b2f412df915fc36ee576bf291ccf472ca657343

Observation 26572b08-138c-4f2f-8822-ccf26c78472f · outbound

This paper cites MOS vs . AB: evaluating text-to-speech systems reliably using cluster ed stan- dard errors,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis MOS vs . AB: evaluating text-to-speech systems reliably using cluster ed stan- dard errors,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.015992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.709252Z digest=sha256:f812434396e420833939918053a1b719b63db0d701e7d666c4750e1db7c5fad7

Observation 422c058a-9164-4776-a0b1-ed510877abc7 · outbound

This paper cites V oxSim: A perceptual voice similarity da taset,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V oxSim: A perceptual voice similarity da taset,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.835953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.712118Z digest=sha256:6c1cea3d7fd17a508d56a651192b57cc789c46428cafd8d2b2c8578d7cfcb88b

Observation e7fd25d3-fd91-4c7c-832e-14a34f7cd461 · outbound

This paper cites Salmon: A suite for acous tic language model evaluation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Salmon: A suite for acous tic language model evaluation,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.663293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.751685Z digest=sha256:93b133f034d6f32cc129e9d676aa64f932b7814c40ba15c67724498adc613c8f

Observation 82dfec55-c787-4bd5-b723-5017a25bfc68 · outbound

This paper cites Automatic evaluation of speaker simila rity,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Automatic evaluation of speaker simila rity,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.472116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:14.887907Z digest=sha256:758015cca4d0eb39497b273edab0ba0155708b393c74fd61ed96a45b24c4b546

Observation 36a86cc9-028d-4d05-b97d-d2dfadd18548 · outbound

This paper cites SVSNet: An end-to-end speaker voice simi larity assessment model,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis SVSNet: An end-to-end speaker voice simi larity assessment model,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.297397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.048606Z digest=sha256:9da7f4dc116209c6cccd8fa8ecc02f4e4f9c4e5a8d5b9858718ebada74b6bb10

Observation 502ce3f8-0886-4a72-9e6a-87181406f5a4 · outbound

This paper cites The V oxCeleb Speaker Recognition Challe nge: a retrospective,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The V oxCeleb Speaker Recognition Challe nge: a retrospective,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.152485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.178145Z digest=sha256:367b09d9fb4173818662c1537ca5635098b2bf3870b4d8811b5d46020aa6fa0a

Observation d98022c9-0b21-4b97-a133-f6480c3df776 · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis SpeechBrain: A General-Purpose Speech Toolkit

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:15.310440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:15.310440Z digest=sha256:8d471c8560c58c47c369f0d1a4ef620e4b01a2c290964e398d4ebd6686b59ba2

Observation b2034870-5e78-47c0-8f37-9038336a86b6 · outbound

This paper cites WavLM: large-scale self-supervised pr e-training for full stack speech processing,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis WavLM: large-scale self-supervised pr e-training for full stack speech processing,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.995774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.439938Z digest=sha256:fc21b81b89d58198ebe5a4efd2b862963fc214a5c51f14f97892e6bf5b19fd5f

Observation 65b8ea99-700f-4336-a54f-b31272b74d15 · outbound

This paper cites The CMU Arctic speech databa ses,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The CMU Arctic speech databa ses,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.842760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.547259Z digest=sha256:75b425397051a85802da6557f3efb00814f6e599d19e5dbeb3141d43446c53f8

Observation 56a9ff04-0491-43ae-bee3-ac8a523caf54 · outbound

This paper cites L2-ARCTIC: a non-native english speech c orpus,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis L2-ARCTIC: a non-native english speech c orpus,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.681892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.631101Z digest=sha256:f578e3d7c1ae16cfc0a6c189432f8c40fb80d6586b900ae0173045fdc9ff4092

Observation 33f54206-915c-4f9a-b19d-ddd85c8d66db · outbound

This paper cites Lib- riSpeech: an ASR corpus based on public domain audio books,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Lib- riSpeech: an ASR corpus based on public domain audio books,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.533289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.701551Z digest=sha256:0217b833a44638a272e6023a81a8f21f44d739ca96671d94e90348a1254a1429

Observation e75cb6d1-b3e3-4a99-93e8-8046229d5370 · outbound

This paper cites Analysis of fun- damental frequency, jitter, shimmer and vocal intensity in chil- dren with phonological disorders,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analysis of fun- damental frequency, jitter, shimmer and vocal intensity in chil- dren with phonological disorders,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.376682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.768394Z digest=sha256:c5a2f0e6431e8f4d8ce8eff2400a97e71a63d7c85c67c2b9dba897d985836ba1

Observation af70a177-120a-4554-bcd0-0fe20b521ab6 · outbound

This paper cites V ocal acoust ic analy- sis – jitter, shimmer and HNR parameters,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V ocal acoust ic analy- sis – jitter, shimmer and HNR parameters,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.199832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.849729Z digest=sha256:cfa33da34d42b166dc68c9148ffb160529f1a1495265099086db9c4c60bc2aca

Observation d80aa272-c9b8-486a-88f8-28b67beccaef · outbound

This paper cites V ariation of the acoustic parameters: f0, jitter, shimmer and alpha ratio in relation with different background noise levels,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V ariation of the acoustic parameters: f0, jitter, shimmer and alpha ratio in relation with different background noise levels,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.053593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:15.945168Z digest=sha256:1dde308507cef305e0b754428736ca7993bbbc6c833e9f2696cab86b8fe11cf5

Observation 6bc0d3b7-7283-4f48-948d-5ea32a83db7b · outbound

This paper cites Rhyth m modeling for voice conversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Rhyth m modeling for voice conversion,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.800943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:16.041293Z digest=sha256:d449642b3461a5b60c18878cb1ff077c2a2da28e470708dc665d1f0328e71709

Observation f993c02a-30c5-4bb9-832e-82ceaf2ffa39 · outbound

This paper cites Opensmile: the munich versatile and fast open-source audio feature extractor,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Opensmile: the munich versatile and fast open-source audio feature extractor,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.468730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:16.162090Z digest=sha256:5f8bb702612ac2181a72f47f8be91e293cb09300ae47b38e23da0fbec65b4a26

Observation 35d8cdf7-a228-477d-8e01-9ae7453031b1 · outbound

This paper cites ASRRL-TTS: Agile Speaker Representation Reinforcement Learning for Text-to-Speech Speaker Adaptation.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ASRRL-TTS: Agile Speaker Representation Reinforcement Learning for Text-to-Speech Speaker Adaptation

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:16.704034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:16.321959Z digest=sha256:bcc67551fadf4b82c34f2965eb108235f6b0c0cab795aa144156b7b8418ece22

Observation fd0d3d8b-c6c7-42cb-afba-aeecf2eaf8f8 · outbound

This paper cites All about audio equalization: Solu- tions and frontiers,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis All about audio equalization: Solu- tions and frontiers,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.192615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:16.437239Z digest=sha256:6f80408bfca752e7332ae8ea55dfcff8f870ac0ac03a8f07a61df3779aa10aba

Pith citing papers

Observation 517abf04-5934-49ef-95b2-e23ab2515c83 · inbound

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis cites this paper.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:17.833571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:40:10.960719Z digest=sha256:38d6c10e65aeeef8e404bc454946d097f97d795c4851f0e21dad1eba3db7d0ed