Pith. sign in

Paper Citation Record · LEDGER

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

As of 10 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 1 inbound Pith citation observation for arXiv:2507.02176.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02176 v1

Coverage vector

measured 69 of 69 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:40:16.437239Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:40:10.960719Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T20:40:17.636566Z

Reference resolution

69 of 69 outbound references displayed

  • verified exact4
  • verified fuzzy56
  • unresolved8
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fbd3ef56-0b7a-4d69-8241-d50fb04b296f · outbound

This paper cites Virtual chara cters are expected to possess unique identities that remain consi stent across time.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Virtual chara cters are expected to possess unique identities that remain consi stent across time

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:29.587245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:10.892442Z digest=sha256:99dde1bc98b4629535a8265cb7fda039067f22f694d89a579f022ae9f5c931fc

Observation 517abf04-5934-49ef-95b2-e23ab2515c83 · outbound

This paper cites Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:17.833571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:10.960719Z digest=sha256:f58ac9a1de358b72ff8265e98a10608d629bbcbeb7260384640dd1d46ebe7e29

Observation f1f5725e-124c-47c6-963b-359f243ac434 · outbound

This paper cites We explore which mark- ers are represented in some widely used ASV embeddings, and measure the effect of confounding factors.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We explore which mark- ers are represented in some widely used ASV embeddings, and measure the effect of confounding factors

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.996348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.033440Z digest=sha256:62a64e0ea3fd1c98ba0c04e61271e6498ce064e22e8a9a1b985b04b397003ed7

Observation ce4bbb54-3138-4eec-b835-4404f612eb6d · outbound

This paper cites This reflects the lower bound of our metric, where we expect the smallest distances.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis This reflects the lower bound of our metric, where we expect the smallest distances

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.311331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.119053Z digest=sha256:b28a92562b0f55b973d894636987f6c514c6b334c011023bcbfe664de285e97e

Observation 8bbfef3e-0a69-429e-9162-8340ecfa3475 · outbound

This paper cites This setting tests wh ether our metric can distinguish between speakers undistinguish - able with speech rate.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis This setting tests wh ether our metric can distinguish between speakers undistinguish - able with speech rate

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:28.102856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.225540Z digest=sha256:d1aa4b4175892c1bb40b8c45e09e13c90096c80234db6104174842871f17356b

Observation 89f98804-a7b9-49a1-99e0-d0bacc9f6e2c · outbound

This paper cites The results in the top section of Table 3 show that rhythm distances are significantly larger between different speak ers, even those with similar speech rates.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The results in the top section of Table 3 show that rhythm distances are significantly larger between different speak ers, even those with similar speech rates

Reference 6

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:40:27.992700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.368279Z digest=sha256:efb5a66c9c5f31d08816f66a6993db83258368d9251223c2b5d51b709445a8dc

Observation 0ec6aae4-9475-4471-a80b-ff699a3f4477 · outbound

This paper cites We showed that ASV embeddings mainly encode speech identity markers relating to anatomy (e.g.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We showed that ASV embeddings mainly encode speech identity markers relating to anatomy (e.g

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.880864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.464887Z digest=sha256:62a7f432cb09b8a379c3c81259c9816836e8540c69166d701bf8184bfb5f7112

Observation 5d5d9667-5627-4180-8746-56411436c248 · outbound

This paper cites An image is worth one word: Personalizing t ext-to- image generation using textual inversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis An image is worth one word: Personalizing t ext-to- image generation using textual inversion,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.756865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.537350Z digest=sha256:a1eb1bfff4720224ec66306e65da0bb9c8a72c154e1d0030effc5e587c7fa11f

Observation 932e5017-ad30-4164-a120-a6d46a36daa7 · outbound

This paper cites Dreambooth: Fine tuning text-to-image d iffusion models for subject-driven generation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Dreambooth: Fine tuning text-to-image d iffusion models for subject-driven generation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.588974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.595403Z digest=sha256:59759629c33d2e2c3789c0ab5aebeff152a0e68e1625ebd085a4a4dcaa3e62f4

Observation 0c62ff45-0144-4f18-aa42-7e8b96b9b4f1 · outbound

This paper cites ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:11.679021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:11.679021Z digest=sha256:da6991931092f7476d1156ddb2d7baeae64380485f894713ea456ee364390dbd

Observation c89f5a9f-9e70-4ff9-a821-38f5722804a4 · outbound

This paper cites CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:11.744848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:11.744848Z digest=sha256:abc30cfd59331ced0c9848e2f34f7acc71b5d46176901d5205e5316b78f7a5e8

Observation 5d17b08b-2cef-42f0-b20c-065a1b807fe3 · outbound

This paper cites Generat ing video game scripts with style,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generat ing video game scripts with style,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.458551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.852395Z digest=sha256:ce5f588970c780f12b03b7c12c7c4dae2afeedb8f9033fc22d0fd960b4b79775

Observation 988925a3-6791-4cc9-b3ea-6d3176cae942 · outbound

This paper cites Meet your favorite character: Open-domai n chat- bot mimicking fictional characters with only a few utterance s,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Meet your favorite character: Open-domai n chat- bot mimicking fictional characters with only a few utterance s,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.351960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.916503Z digest=sha256:b874109bf282f68ccfe1d5a753a11cce8f2ef83680e8ee455bcf7e2e219b1e8b

Observation ba599d21-aa4a-44e9-9de5-a9c72b8282d5 · outbound

This paper cites Information conveyed by vow- els.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Information conveyed by vow- els

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:27.170164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:11.991064Z digest=sha256:3ccfeb119440d64f84ddc9cb44f217b77b0094a67cc1668d8a0f00f78e11fe3b

Observation d6d6a64a-a140-4335-95bd-c7a4a8d02172 · outbound

This paper cites The perception of personal iden tity in speech: Evidence from the perception of twins’ speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The perception of personal iden tity in speech: Evidence from the perception of twins’ speech,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.987242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.078389Z digest=sha256:edd18e031193eb583f2956ee97c7a044ff536d8ab816293c49e21bd542b46e06

Observation e7c45ab1-d99e-4c5a-812e-3435e2d3b198 · outbound

This paper cites X-V ectors: Robust DNN embeddings for speaker recognition,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis X-V ectors: Robust DNN embeddings for speaker recognition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.806842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.158681Z digest=sha256:64b17f932d70f067c64f2bb65417719edad1661d4e3df953c4738654bb2d7b91

Observation f7af280c-6e79-496f-a69c-c5d8babb5c99 · outbound

This paper cites ECAPA - TDNN: emphasized channel attention, propagation and aggre ga- tion in TDNN based speaker verification,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ECAPA - TDNN: emphasized channel attention, propagation and aggre ga- tion in TDNN based speaker verification,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.629537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.245018Z digest=sha256:6ad2eb780a63f5c88b27d9a6bb1ea75df4cdb6c4229723d363583591611eb2f8

Observation cc153765-41c9-49ce-b995-37e9c0f1de4d · outbound

This paper cites State-of-the-art speaker recogni tion with neural network embeddings in NIST SRE18 and speakers in the wild evaluations,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis State-of-the-art speaker recogni tion with neural network embeddings in NIST SRE18 and speakers in the wild evaluations,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.482722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.319383Z digest=sha256:ec34cac3d071b462308a3dc651b4910d7f95b3dd731ddb221587147645e68b48

Observation cb4977ac-8be1-44ea-ac18-7e54da19fefd · outbound

This paper cites Generalized end-to-end loss for speaker verifica- tion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generalized end-to-end loss for speaker verifica- tion,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.302718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.410015Z digest=sha256:076c45d9434330701a1df78a4b9566a454396ee29fbc77163472388de68a66d4

Observation b5ee12e4-44fe-4f82-9642-408fa023c43c · outbound

This paper cites We need variations in speech synthe- sis: Sub-center modelling for speaker embeddings,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis We need variations in speech synthe- sis: Sub-center modelling for speaker embeddings,

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-06T20:40:17.374269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.504831Z digest=sha256:f8491238885cf937eb7778664c254f1628ed8d5dc5003b7abde29a2510ffeb09

Observation 7f5eebdb-5db5-4fd2-87f8-7537d1edc768 · outbound

This paper cites Predictions of subjective ratings and spoofing as- sessments of V oice Conversion Challenge 2020 submissions,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Predictions of subjective ratings and spoofing as- sessments of V oice Conversion Challenge 2020 submissions,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:26.121033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.603592Z digest=sha256:eaa5a42b8ff53809c0539b7fb1efb05c7dc915168a11f75fdf9c7727ee872aca

Observation b0d89eba-6eef-4933-97ed-e2d568c16f86 · outbound

This paper cites Transfer learning from speaker verificat ion to multi- speaker text-to-speech synthesis,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Transfer learning from speaker verificat ion to multi- speaker text-to-speech synthesis,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.902006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.679830Z digest=sha256:cf230d05c71f30b8ad4fce9384801101bd5452f2b9880b124b4fedf51ae0d31c

Observation d59ce44c-d9f4-412b-bedd-6b5094c18c70 · outbound

This paper cites Zero-shot multi-speaker text-to-sp eech with state-of-the-art neural speaker embeddings,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Zero-shot multi-speaker text-to-sp eech with state-of-the-art neural speaker embeddings,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.728737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.740868Z digest=sha256:f51ea07cbea35a83da3387c841bb9df5a59952a1f9b87734cf2db6030f2e7774

Observation 38425e51-2504-4a84-874e-c042e926c1d6 · outbound

This paper cites Y ourTTS: towards zero-shot multi- speaker tts and zero-shot voice conversion for everyone,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Y ourTTS: towards zero-shot multi- speaker tts and zero-shot voice conversion for everyone,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.600183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.837534Z digest=sha256:c63aa5d9163df24e2cbf70b9fcadc889cb1b29393485434b3841ce026d0b2ac3

Observation 1cb6588a-a323-4f5c-b318-afc26a02eae4 · outbound

This paper cites Investigating on incorporating pret rained and learnable speaker representations for multi-speaker mult i-style text-to-speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Investigating on incorporating pret rained and learnable speaker representations for multi-speaker mult i-style text-to-speech,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.420748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:12.936260Z digest=sha256:5fde464d7dd38eab9297c2c6c6fcc04afe033463722e5e1ca04acb46549415d3

Observation e3bf7555-52cb-4bf9-bd88-e292621223c0 · outbound

This paper cites Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.032052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.032052Z digest=sha256:79c99ea36aedcffc7d0b958ce048a62ae18631effce5a8c985b794bd3564fcec

Observation 01a602c7-7a3a-4992-bcca-2d20ea94f10d · outbound

This paper cites The Multi-Speaker Multi-Style V oice Clo ning Chal- lenge 2021,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Multi-Speaker Multi-Style V oice Clo ning Chal- lenge 2021,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.237363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.128895Z digest=sha256:017c31ef380512765353c75920d0be69315f81fe538affd67640c6e9c08b7ea1

Observation 4857260a-648a-4d2c-80a3-34fd08fb25de · outbound

This paper cites VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.209463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.209463Z digest=sha256:64fe6fcfc69de4270639008d8813c7e3314857b1d0e83e6144aa2cbafaebe216

Observation fc46cf74-cdd9-420c-b7f8-d04b70d048b7 · outbound

This paper cites Evaluating text-to-speech synthesis from a large discrete token-based speech language model,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Evaluating text-to-speech synthesis from a large discrete token-based speech language model,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:25.020012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.330030Z digest=sha256:65cc502e946367c7d331a09cf6d41f44bdbc44656a5d5c9a3981ffb7457db6ae

Observation f6e68c49-0f0f-4721-b3a1-898330ee7271 · outbound

This paper cites A comparison of discrete and soft speech units for improved voice conversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis A comparison of discrete and soft speech units for improved voice conversion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:13.420536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:13.420536Z digest=sha256:600f21661c90d787a7f311e5efe41c72a24607ef3246b0d5adad0996c276b2ba

Observation 9fdce733-cd03-4761-b18d-ef7f05f91428 · outbound

This paper cites V oicebox: Text-guided multilingual univ ersal speech generation at scale,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V oicebox: Text-guided multilingual univ ersal speech generation at scale,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.864555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.484801Z digest=sha256:175cfd8efc5b9bf2e5304a8bb27bb741bb777a57c59a9b0e63acae6ef57d1238

Observation 67b8fbcc-5f6d-447d-bf18-4f153dc8ce3d · outbound

This paper cites Neural codec language models are zero-s hot text to speech synthesizers,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Neural codec language models are zero-s hot text to speech synthesizers,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.567138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.544795Z digest=sha256:266b2d07cde1e0f8c7c69df9853513f7df6f43b7825bf20a0f2646b38536793a

Observation fc5d7913-3f76-4168-90e3-16b024875a87 · outbound

This paper cites The Singing V oice Conversion Chall enge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Singing V oice Conversion Chall enge 2023,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.362172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.610042Z digest=sha256:d2823eee87300c4737ab714b95f946e9ba578d7db92560e9dd95c034114e3ca4

Observation f3ef1777-0b8f-4ab0-9ddc-4580ee3c651c · outbound

This paper cites Generative Data Augmentation Challe nge: Zero- shot speech synthesis for personalized speech enhancement ,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Generative Data Augmentation Challe nge: Zero- shot speech synthesis for personalized speech enhancement ,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:24.160692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.678518Z digest=sha256:2721003bb028fdca78db9a8003eac9c856ac58f74b9c35224722bd594a634edc

Observation 654da8cc-1fdc-4b24-b46c-06147b004493 · outbound

This paper cites Acoustic properties of voice timbre t ypes and their influence on voice classification,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Acoustic properties of voice timbre t ypes and their influence on voice classification,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.920263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.758589Z digest=sha256:8f3f4911b74af0e038942ab90d1f7a1f8e31ceda8381861c5eea994863e3afa1

Observation 5c79d98a-4962-4f10-802e-f487f2340f69 · outbound

This paper cites Speech production patterns in producing l inguis- tic contrasts are partly determined by individual differen ces in anatomy,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Speech production patterns in producing l inguis- tic contrasts are partly determined by individual differen ces in anatomy,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.680062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.823753Z digest=sha256:e61b3acb31ba7aeeaf7b8a1039c1bd621cd60d071d2cc68411c60fb89abdfed5

Observation bc66290e-98e4-4bd5-981a-3190fe79369b · outbound

This paper cites an unresolved cited work.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:40:23.460801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.871818Z digest=sha256:920757422c17351d8d18d0a9305098603027920e4b41718ef6871b7169d0b27d

Observation cae77f0c-21b8-4b0b-8ad9-32e71334c683 · outbound

This paper cites A study of rhythm in London: Is syllable-timing a feature of Multicultural London English ?.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis A study of rhythm in London: Is syllable-timing a feature of Multicultural London English ?

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.289528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.928173Z digest=sha256:ff6a0a3ad3b9b0890a6ad170ab163ea7e1f69cba97d8edb9e4c79e9275dc8b6e

Observation af5d4057-27d1-48f6-961b-011f8958c4ba · outbound

This paper cites The measurement of rhythm: a comparison of Sin- gapore and British English,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The measurement of rhythm: a comparison of Sin- gapore and British English,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:23.067362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:13.978443Z digest=sha256:8a427e08fbb6aa89d68329340f5de0369e2abc6da58bd0f1533a6b73d6d3af58

Observation 10783b19-ca12-4bc5-8432-b0df7d8a26ba · outbound

This paper cites Sociophonetics of phonotactic pheno mena in french,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Sociophonetics of phonotactic pheno mena in french,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.807797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.042470Z digest=sha256:2818b0f814ad390a8e342a435dc77f6e87ed8da63fb1506474668d280c2abee6

Observation 4d90a904-75f6-4cf3-8437-f4ba2f576871 · outbound

This paper cites Individual di fferences in vowel production,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Individual di fferences in vowel production,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.620821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.108230Z digest=sha256:e533aae64eb7eb38b1b29c6c1a80f6efd2e4910dcf5864f19cbab31080c88aae

Observation 13627fac-ceff-4ca7-9de0-37f89d1c47df · outbound

This paper cites Individual differences in speech pro- duction: V oice-onset-time,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Individual differences in speech pro- duction: V oice-onset-time,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.428908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.167152Z digest=sha256:d17ee51fa20659c16f43e51d0f0c067dd6c88fa7d2ba2dabdf23b19021513ce7

Observation 6c2ed39d-60f7-4358-bb49-930065f56271 · outbound

This paper cites The social life of phonetic s and phonology,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The social life of phonetic s and phonology,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.226893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.225970Z digest=sha256:45148b064bd20466bd0baa8aa7f38837765f0c0495ef79afb9c67a3eb06d0581

Observation a0a78c3e-a471-4b32-9051-7867fd9eefd9 · outbound

This paper cites Prosodic rhythm and Afric an American english,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Prosodic rhythm and Afric an American english,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:22.055134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.273503Z digest=sha256:07f36ac865e2002726ccdc8b9bf8575933552e4fcdb04ef4ac459d86fc770fd1

Observation dd9e3107-7a7a-413a-82ef-74ec4916a3e9 · outbound

This paper cites La liaison sans enchaˆ ınement,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis La liaison sans enchaˆ ınement,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.865653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.325252Z digest=sha256:fb7ef0f569dba61a193cbaa1267f73c8bdeccf767b58ae681d30681c3a78bb6c

Observation f7e19809-8f4a-47d6-b320-721f6b2d4a4b · outbound

This paper cites Stuart-Smith, E.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Stuart-Smith, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.707914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.381549Z digest=sha256:17e05e6049c96ff66623e931f1a29eaab6615a59982d5ed7943dbbf5fa5b9024

Observation f2fee590-4f65-4a33-85bb-bb5f1f5ae99f · outbound

This paper cites The Blizzard Challenge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The Blizzard Challenge 2023,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.492072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.437994Z digest=sha256:dc748dfa947f4ab5844729802190f8c148a85f7e62f080ad14393155d61db96e

Observation 7c04876d-3c13-4c25-8171-85a9405c39fb · outbound

This paper cites Refining the evaluation of speech synthesis: A summ ary of the blizzard challenge 2023,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Refining the evaluation of speech synthesis: A summ ary of the blizzard challenge 2023,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.310864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.486042Z digest=sha256:69289a773cfe66270e69a9f55a5df4c5ddc6b0266e66007cad506186214f0140

Observation 457a1f57-16e9-4294-83bd-5d1340707f4c · outbound

This paper cites Good practices for evaluation of synthesized speech,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Good practices for evaluation of synthesized speech,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:14.536826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:14.536826Z digest=sha256:834b7d3ae1d4180f68295bb175b0223dc4311cf051a7be67eef2aed67ed34ae0

Observation a8594ffb-9d77-4924-a681-932f2c5a1c65 · outbound

This paper cites Stuck in the MOS pit: A critical anal ysis of MOS test methodology in TTS evaluation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Stuck in the MOS pit: A critical anal ysis of MOS test methodology in TTS evaluation,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.155580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.599962Z digest=sha256:f9e3fccd58197ae03f3dd761cb34da8b23a0f3adaa88fac65246ed6ee9b2e46e

Observation bf0a2f81-a772-4739-9fd1-56a402a46427 · outbound

This paper cites Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:16.930605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.655226Z digest=sha256:848d8478a866e40d61e1f82b31caf58878d0530a45365828cb2ff7886a53a5c3

Observation 26572b08-138c-4f2f-8822-ccf26c78472f · outbound

This paper cites MOS vs . AB: evaluating text-to-speech systems reliably using cluster ed stan- dard errors,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis MOS vs . AB: evaluating text-to-speech systems reliably using cluster ed stan- dard errors,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:21.015992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.709252Z digest=sha256:f25f61adafb243523e1558531660ede5b1cd1f3f36289a2cfbb4537ff4a3e506

Observation 422c058a-9164-4776-a0b1-ed510877abc7 · outbound

This paper cites V oxSim: A perceptual voice similarity da taset,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V oxSim: A perceptual voice similarity da taset,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.835953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.712118Z digest=sha256:8fab9b502712fbb5e11032ad9a1df94e8bc4336941a06835e59f8fbb1c109ab0

Observation e7fd25d3-fd91-4c7c-832e-14a34f7cd461 · outbound

This paper cites Salmon: A suite for acous tic language model evaluation,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Salmon: A suite for acous tic language model evaluation,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.663293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.751685Z digest=sha256:2a6065d7ca3478d2cabc7a0714db4b6dd6239d4a218975e837484ff265a94546

Observation 82dfec55-c787-4bd5-b723-5017a25bfc68 · outbound

This paper cites Automatic evaluation of speaker simila rity,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Automatic evaluation of speaker simila rity,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.472116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:14.887907Z digest=sha256:709b2d8a3cb965752161c666e6d13030f58234688899d7ac657098a81ad5598e

Observation 36a86cc9-028d-4d05-b97d-d2dfadd18548 · outbound

This paper cites SVSNet: An end-to-end speaker voice simi larity assessment model,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis SVSNet: An end-to-end speaker voice simi larity assessment model,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.297397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.048606Z digest=sha256:e7a5b1d7a9a20f65b57a0dd899a894916c158f0fe4bdf31540bf931e70a2eea4

Observation 502ce3f8-0886-4a72-9e6a-87181406f5a4 · outbound

This paper cites The V oxCeleb Speaker Recognition Challe nge: a retrospective,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The V oxCeleb Speaker Recognition Challe nge: a retrospective,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:20.152485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.178145Z digest=sha256:a7589aaf54ed8a91ac10431e7d38f764a09a7bd4669d62c01dd816feebb1c39b

Observation d98022c9-0b21-4b97-a133-f6480c3df776 · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis SpeechBrain: A General-Purpose Speech Toolkit

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:15.310440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:15.310440Z digest=sha256:8d471c8560c58c47c369f0d1a4ef620e4b01a2c290964e398d4ebd6686b59ba2

Observation b2034870-5e78-47c0-8f37-9038336a86b6 · outbound

This paper cites WavLM: large-scale self-supervised pr e-training for full stack speech processing,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis WavLM: large-scale self-supervised pr e-training for full stack speech processing,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.995774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.439938Z digest=sha256:fd26d838155c6b3e840875febb62cd78ed3e4d257cdc21eda4fe1f60de6fb487

Observation 65b8ea99-700f-4336-a54f-b31272b74d15 · outbound

This paper cites The CMU Arctic speech databa ses,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis The CMU Arctic speech databa ses,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.842760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.547259Z digest=sha256:f2239a4690703ffd351dea4560b335eea08dd2314294766954fcb75594f041d5

Observation 56a9ff04-0491-43ae-bee3-ac8a523caf54 · outbound

This paper cites L2-ARCTIC: a non-native english speech c orpus,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis L2-ARCTIC: a non-native english speech c orpus,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.681892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.631101Z digest=sha256:2045455a68fbea6a5007a1f05c0314a314cd552fe9b6790ee564a5a4461214ac

Observation 33f54206-915c-4f9a-b19d-ddd85c8d66db · outbound

This paper cites Lib- riSpeech: an ASR corpus based on public domain audio books,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Lib- riSpeech: an ASR corpus based on public domain audio books,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.533289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.701551Z digest=sha256:aaf7a8c783977f5c70f278c803e4d7bad1400f1525a671b79f7319e01d7b895d

Observation e75cb6d1-b3e3-4a99-93e8-8046229d5370 · outbound

This paper cites Analysis of fun- damental frequency, jitter, shimmer and vocal intensity in chil- dren with phonological disorders,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analysis of fun- damental frequency, jitter, shimmer and vocal intensity in chil- dren with phonological disorders,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.376682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.768394Z digest=sha256:f589797f1df7d0065523dc0ffb81f1276300a8306ddef15b870c092d2dec1478

Observation af70a177-120a-4554-bcd0-0fe20b521ab6 · outbound

This paper cites V ocal acoust ic analy- sis – jitter, shimmer and HNR parameters,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V ocal acoust ic analy- sis – jitter, shimmer and HNR parameters,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.199832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.849729Z digest=sha256:f3ff721b1c31e672a1efaa77906c63e4edfa9ae94727d9af7db6464bf2730f7c

Observation d80aa272-c9b8-486a-88f8-28b67beccaef · outbound

This paper cites V ariation of the acoustic parameters: f0, jitter, shimmer and alpha ratio in relation with different background noise levels,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis V ariation of the acoustic parameters: f0, jitter, shimmer and alpha ratio in relation with different background noise levels,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:19.053593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:15.945168Z digest=sha256:92014123a67e498cfc9e1d04f4113e8f876a4407c10dc65d648fb7cb50dd51d7

Observation 6bc0d3b7-7283-4f48-948d-5ea32a83db7b · outbound

This paper cites Rhyth m modeling for voice conversion,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Rhyth m modeling for voice conversion,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.800943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:16.041293Z digest=sha256:68b593fa28e24203fa94c2000e32d84531ea5264b973ec4236356c8ff5c3202d

Observation f993c02a-30c5-4bb9-832e-82ceaf2ffa39 · outbound

This paper cites Opensmile: the munich versatile and fast open-source audio feature extractor,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Opensmile: the munich versatile and fast open-source audio feature extractor,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.468730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:16.162090Z digest=sha256:d68bc8561a2a17d09e4e387d7a8d5987ef2711d1f3154c754634818fbca541f9

Observation 35d8cdf7-a228-477d-8e01-9ae7453031b1 · outbound

This paper cites ASRRL-TTS: Agile Speaker Representation Reinforcement Learning for Text-to-Speech Speaker Adaptation.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis ASRRL-TTS: Agile Speaker Representation Reinforcement Learning for Text-to-Speech Speaker Adaptation

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:16.704034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:16.321959Z digest=sha256:52246fab38cdc1066139580f55b566f90079945eec33d6e13c0931231ca4810f

Observation fd0d3d8b-c6c7-42cb-afba-aeecf2eaf8f8 · outbound

This paper cites All about audio equalization: Solu- tions and frontiers,.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis All about audio equalization: Solu- tions and frontiers,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:40:18.192615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:16.437239Z digest=sha256:c02c2f85a6cb69eeb9a3623629cb37d923dddcdef0e7c931ffff96b27c0a4601

Pith citing papers

Observation 517abf04-5934-49ef-95b2-e23ab2515c83 · inbound

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis cites this paper.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:17.833571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:40:10.960719Z digest=sha256:f58ac9a1de358b72ff8265e98a10608d629bbcbeb7260384640dd1d46ebe7e29