Pith. sign in

Paper Citation Record · LEDGER

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation

As of 15 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 2 inbound Pith citation observations for arXiv:2411.12719.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.12719 v3

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:18:09.405163Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:54:25.134302Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T20:40:16.820041Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact7
  • verified fuzzy19
  • unresolved16
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3c1a1d6f-6ab1-4f5a-bfce-5003d081fa0d · outbound

This paper cites Using vaes and normalizing flows for one-shot text-to-speech synthesis of expressive speech.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Using vaes and normalizing flows for one-shot text-to-speech synthesis of expressive speech

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.148265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.148265Z digest=sha256:78e530ab8bb2d0f76ae6df0de8c82f1d4fe0f3dcf8c9423ad368d25dabdc43d5

Observation 0aa6f9ed-132a-48ef-8506-a28af09800f1 · outbound

This paper cites an unresolved cited work.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:18:10.555075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.154175Z digest=sha256:39c44d4ebc1076595d68c66471320db3855b67462081c20590360d4a7105a1ca

Observation 8c0b2777-7990-482a-8374-dd597e78d82a · outbound

This paper cites The use of rating and likert scales in natural language generation human evaluation tasks: A review and some recommendations.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation The use of rating and likert scales in natural language generation human evaluation tasks: A review and some recommendations

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.159380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.159380Z digest=sha256:354147725947c746f290a7e260c41798f7badece51b9c2aee4b4a5a06d0cbc06

Observation f2c6b891-4396-4e63-9d34-1f852f27a95a · outbound

This paper cites Resources for indian languages.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Resources for indian languages

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.537451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.167050Z digest=sha256:8e68eb6b5eb80a0cd6ae89afc916ef09245420f3667d304f351006615ce9a9de

Observation 3bfeacaf-c816-4c98-9c32-32527f345816 · outbound

This paper cites an unresolved cited work.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:18:10.519058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.173603Z digest=sha256:d6f874aa703a3a5670c511da48df23f47ba71f56a564120e2d85d40d18f875c4

Observation 0798350d-df6c-4edf-a55c-eb7011dad3ed · outbound

This paper cites Why we should report the details in subjective evaluation of tts more rigorously, 2023.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Why we should report the details in subjective evaluation of tts more rigorously, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.496888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.179024Z digest=sha256:ca1bb5d2a0543a443fc54317d1e44896797c2b9a29381207a19a1a201caf6f4d

Observation 20e7ac0f-1155-41ec-9dfb-92156395ba87 · outbound

This paper cites Evaluating Long-form Text-to-Speech: Comparing the Ratings of Sentences and Paragraphs.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Evaluating Long-form Text-to-Speech: Comparing the Ratings of Sentences and Paragraphs

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:18:09.897053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.185526Z digest=sha256:d5417d6e9b55fdfa6aba5663ece0c95702ad083e7f952453bf389c6efdfa2966

Observation bc4de7d8-c9a2-405f-b6fa-e84bcd40d1a3 · outbound

This paper cites Investigating Range-Equalizing Bias in Mean Opinion Score Ratings of Synthesized Speech.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Investigating Range-Equalizing Bias in Mean Opinion Score Ratings of Synthesized Speech

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.192725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.192725Z digest=sha256:7b55ad66ec453efb9cd7a46d3a7945ce6f79846b8ad0bb8637901071fe3ab5c0

Observation 3d2b61d5-d69c-4f51-9b7e-c848f17d3369 · outbound

This paper cites an unresolved cited work.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:18:10.480559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.199118Z digest=sha256:9471ab4aaf0924af88009d741c1a46b5607eef882ff3e7aff3cb6bbb2a5758b0

Observation 6056c3e1-ac68-4e36-ba1b-e744208ee13e · outbound

This paper cites The authenticity gap in human evaluation.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation The authenticity gap in human evaluation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.208026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.208026Z digest=sha256:30db53ccca7feb204898ee9586822d8fe50ff88408849a0aae138c04fc9f4e0e

Observation 9b9855c3-beef-484b-9da7-4e33492f2738 · outbound

This paper cites Importance of human factors in text-to-speech evaluations.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Importance of human factors in text-to-speech evaluations

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.466522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.212755Z digest=sha256:dd626557f84caaa52a6c6d64ea3cc249df6b1cb8d251ca812a56bc52fab6c312

Observation 488f3f03-aa9c-4a34-8549-3566fbacc2f9 · outbound

This paper cites Investigating the robustness of sequence-to-sequence text-to-speech models to imperfectly-transcribed training data.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Investigating the robustness of sequence-to-sequence text-to-speech models to imperfectly-transcribed training data

Reference 12

Resolution
verified exact
doi, observed 2026-08-12T17:18:09.610866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.217603Z digest=sha256:ca7243d480b386cf18ff5a5d7a58ec3068ab6908e96b83d7d57398c1608a3f83

Observation 0865dbf4-d125-4acf-832d-9b55063bb0a2 · outbound

This paper cites Foster, David Grangier, Viresh Ratnakar, Qijun Tan, and Wolfgang Macherey.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Foster, David Grangier, Viresh Ratnakar, Qijun Tan, and Wolfgang Macherey

Reference 13

Resolution
malformed identifier
no resolver link, observed 2026-08-12T17:18:09.226557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.226557Z digest=sha256:50b5e693ed67ccfed6a96f6139256d164cf7b54f1917151b84060946fbd618a0

Observation 55432e4d-c366-47d7-9ceb-bd9ed04b4936 · outbound

This paper cites Howcroft and Verena Rieser.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Howcroft and Verena Rieser

Reference 14

Resolution
verified exact
doi, observed 2026-08-12T17:18:10.451433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.231480Z digest=sha256:a7dbb26efba161d15a5fe6e5d1b49ab0191b57fd00841cf914efb9f50ab55aeb

Observation ad3f8f47-9fad-4a54-aa3c-d4a9ac3b1fd7 · outbound

This paper cites Low-resource expressive text-to-speech using data augmentation.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Low-resource expressive text-to-speech using data augmentation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.437259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.238845Z digest=sha256:1b245fa9d50faf8896a8e01fac23636048eb3dbf36241ba812de4dc4c4a8e63b

Observation 9af76b59-9ace-4aa2-87c4-efcc74402031 · outbound

This paper cites Method for the subjective assessment of intermediate quality level of audio systems.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Method for the subjective assessment of intermediate quality level of audio systems

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.421236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.246677Z digest=sha256:cca0d12e5bb274967e81aed2e31ac05fe33f9dd0866daf07f41b5f7ec961f9b6

Observation 78462dcb-9221-4818-80f5-e25112270ce5 · outbound

This paper cites Naturalspeech 3: Zero-shot speech synthesis with factorized codec and diffusion models, 2024.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Naturalspeech 3: Zero-shot speech synthesis with factorized codec and diffusion models, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.398580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.255779Z digest=sha256:3b34c0828277ce053f10f5497a36c8ee197601094443dce1c4f536de880d3a12

Observation 3f56a0d5-1c3f-4f4d-9df7-27b28c14e2b9 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.375565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.261227Z digest=sha256:e6237ae61726dc45355d3395f7e7a16f165933b17c6043b0ab2ce7869b96c438

Observation 51e42834-d704-4d2a-b5e2-8fae5f0882a4 · outbound

This paper cites Stuck in the MOS pit: A critical analysis of MOS test methodology in TTS evaluation.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Stuck in the MOS pit: A critical analysis of MOS test methodology in TTS evaluation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.267333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.267333Z digest=sha256:05d240e67b3f1894df3b2b067bf76973bde9e8012056e45b3fd82167e167478e

Observation 27a62373-f535-43f2-82ed-28128d7336b7 · outbound

This paper cites On the stability of system rankings at WMT.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation On the stability of system rankings at WMT

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.353124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.282393Z digest=sha256:5c76aff21c087745c5027b79e61563694da6ebf9c5a278968e320ce241289f39

Observation 101bb44d-74ce-4dbe-8fc9-3a6fca200f6b · outbound

This paper cites Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis, 2020.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis, 2020

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.340182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.287288Z digest=sha256:60fd8fcab01c48609969cda19b0d510e745f4639491a65102a7eb32db381556f

Observation 3aa214a7-92c9-4c21-9aae-d201276b5b50 · outbound

This paper cites Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.326004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.292188Z digest=sha256:56f3ddc57fef43010de9bf3aaf8ba1c3b92502bfaffb6a354d164e554803ae18

Observation 125bf691-a08f-4ab2-b416-d53073e7e153 · outbound

This paper cites Khapra, and Karthik Nandakumar.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Khapra, and Karthik Nandakumar

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.297324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.297324Z digest=sha256:a8420438429c8aa19ebf8c4ffd90fbdaf4f349d9ef5c526ddcd941d2d3093c1b

Observation 2109b0fb-e70f-4532-9c97-c72a5b98e7e0 · outbound

This paper cites BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.304007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.304007Z digest=sha256:4883624d8ae2907c9134bd27b1eb6abf131063321f5c58ee12dfc29f471cdf30

Observation 95c49a17-b981-4e98-abc3-13a5ed164972 · outbound

This paper cites Fastpitch: Parallel text-to-speech with pitch prediction.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Fastpitch: Parallel text-to-speech with pitch prediction

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.311160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.310421Z digest=sha256:021c2ddbb58856b89228ddee66160bf4581a54a992cee3aa219a9e1db577948e

Observation 2d500fa2-01e3-4657-9645-c87e5c51fb10 · outbound

This paper cites The limits of the mean opinion score for speech synthesis evaluation.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation The limits of the mean opinion score for speech synthesis evaluation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.314684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.314684Z digest=sha256:adc39ee023d0ae28d3da7bdb19d4a4fc363a3bd6ca985354c937b943d545d984

Observation d5edfebc-c5e1-416d-933d-ce58fe5e30a2 · outbound

This paper cites Style TTS 2: Towards human-level text-to-speech through style diffusion and adversarial training with large speech language models.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Style TTS 2: Towards human-level text-to-speech through style diffusion and adversarial training with large speech language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.294741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.318917Z digest=sha256:727c0800f56f78c1efc2571cfae8c95a6e790bbedd74c500bb218ff37f61818e

Observation ea2ea477-8b3b-4e8c-b724-3f6846797160 · outbound

This paper cites an unresolved cited work.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Unresolved cited work

Reference 28

Resolution
verified exact
doi, observed 2026-08-12T17:18:09.556185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.324258Z digest=sha256:8e454bead2fecfd8b30d930835d30e934c0049611c06d2ed82b846de67d4e1fa

Observation 84c8c95c-2802-4c1a-abd2-f6e26d79a548 · outbound

This paper cites Comprehensive evaluation of statistical speech waveform synthesis.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Comprehensive evaluation of statistical speech waveform synthesis

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.274539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.328635Z digest=sha256:879e78034e871f106e454d9d126bbaf48c37c70bf56bd9ff14b0c767e0590b17

Observation 81823d3d-32d9-40ff-9749-3bc49120092b · outbound

This paper cites Text-free non-parallel many-to-many voice conversion using normalising flows.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Text-free non-parallel many-to-many voice conversion using normalising flows

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:18:09.540836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.336150Z digest=sha256:b4d495b647eef71cb134cf4cf5d60e3eadbd0a8bb0cc1bda73f9db7eaee81681

Observation e347b7c8-ba44-4236-95dd-d5eaf5bdbbfd · outbound

This paper cites Pauwels, S.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Pauwels, S

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.258349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.341545Z digest=sha256:ea25f0d5a636ff87adda6e4c3be880e68c4c6c3543001068f197090894fc4f9a

Observation e46c07cb-0dfc-463a-ac37-21673ec95a2d · outbound

This paper cites Fastspeech 2: Fast and high-quality end-to-end text to speech.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Fastspeech 2: Fast and high-quality end-to-end text to speech

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.241999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.347107Z digest=sha256:a90d181fc3e73a418af65bcbfc00ccf7d7c85cfb777033c3dbd2584763b97385

Observation fe24b8bb-b612-4994-8c48-6fbfbf1389ad · outbound

This paper cites A perceptual investigation of wavelet-based decomposition of f0 for text-to-speech synthesis.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation A perceptual investigation of wavelet-based decomposition of f0 for text-to-speech synthesis

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.225030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.351729Z digest=sha256:76199785d63212c477dae4c6849375c608c422764a7f5df343ad12631f1109b1

Observation c268e341-e10a-4115-973d-dc0acd000a64 · outbound

This paper cites Schoeffler, S.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Schoeffler, S

Reference 34

Resolution
verified exact
doi, observed 2026-08-12T17:18:09.519365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.357581Z digest=sha256:3d17d1b780b5c10e07b26cb5441e04a0c17d085988488bb6d11adc7b99dd2ba7

Observation 7d63e0f3-9a48-4681-ab2e-7a54d21bd07f · outbound

This paper cites Naturalspeech 2: Latent diffusion models are natural and zero-shot speech and singing synthesizers.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Naturalspeech 2: Latent diffusion models are natural and zero-shot speech and singing synthesizers

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.211748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.362000Z digest=sha256:4a3e7b662459136da488f0293788a690d65a689d9bb9f73286f189161df22a35

Observation 369c68ab-77ce-4977-b555-763c75945825 · outbound

This paper cites Better replacement for tts naturalness evaluation.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Better replacement for tts naturalness evaluation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.043581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.367151Z digest=sha256:01cf05cffc60013e63d09c8ab04896bea8a2288c4e6b68e57d842ea30a7527d7

Observation 9a8d8de0-2e9c-4680-9628-8474dd766ae6 · outbound

This paper cites Naturalspeech: End-to-end text-to-speech synthesis with human-level quality.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Naturalspeech: End-to-end text-to-speech synthesis with human-level quality

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.372968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.372968Z digest=sha256:3c18360396767017b2fe272fcfed4909ff30bfcf60b101dc914df8fb590a26df

Observation 6cdb4d4f-dba2-4f3f-8481-45312ffa3997 · outbound

This paper cites Enhancing sequence-to-sequence text-to-speech with morphology.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Enhancing sequence-to-sequence text-to-speech with morphology

Reference 38

Resolution
verified exact
doi, observed 2026-08-12T17:18:09.499428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.378485Z digest=sha256:f561eea905232a4d43841a86a391ffb6040283db077b43b51872b9ae34f26a94

Observation 9c250bc3-f1c8-4f0e-bce2-9ce2b445082c · outbound

This paper cites an unresolved cited work.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:18:10.027653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.383741Z digest=sha256:62f225366a867c48e18456bc9940bb00e68bfcc4a7ca9e8c97aeed357361aa08

Observation a9e50d5d-d77f-4b40-a7a6-a6e4a73f8c3f · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.390006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.390006Z digest=sha256:c6bcbce0f79f57590366846253231339cbd2955ed145b73f98935984d3944832

Observation f9dac723-af04-4971-b7c6-bffeac4e8c5b · outbound

This paper cites Are we using enough listeners? no! - an empirically-supported critique of interspeech 2014 TTS evaluations.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation Are we using enough listeners? no! - an empirically-supported critique of interspeech 2014 TTS evaluations

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.395130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.395130Z digest=sha256:3f7fa5e8d5a42b847b2d05e3e6e649d02ecd5199efe89671f24e6ec8ef624dba

Observation d65f231e-932c-44f3-abd5-85e4ce777921 · outbound

This paper cites On some biases encountered in modern audio quality listening tests (part 2): Selected graphical examples and discussion.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation On some biases encountered in modern audio quality listening tests (part 2): Selected graphical examples and discussion

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:18:10.013218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-12T17:18:09.400092Z digest=sha256:9f1ef0ff41552bf7adb289cd9a570da4727e91ecad591341d9537c4bf290749b

Observation 8904bb61-890b-4dbc-a991-35d5ea78707b · outbound

This paper cites write newline.

Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation write newline

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T17:18:09.405163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:18:09.405163Z digest=sha256:712f19921632748ceb7e9332b9d93cedb0342aca1ce8440e30243834bdc2d829

Pith citing papers

Observation f175090e-0a47-49ab-9984-486cfd83029a · inbound

Phir Hera Fairy: An English Fairytaler is a Strong Faker of Fluent Speech in Low-Resource Indian Languages cites this paper.

Phir Hera Fairy: An English Fairytaler is a Strong Faker of Fluent Speech in Low-Resource Indian Languages Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:25.134302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:25.134302Z digest=sha256:eb0d70310d8c834544139a75e21266146e5ed0c0ddfe8820ccf423d744544d51

Observation bf0a2f81-a772-4739-9fd1-56a402a46427 · inbound

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis cites this paper.

Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:40:16.930605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T20:40:14.655226Z digest=sha256:2e742f20bb80b3bb909979bc57419cb37bdc58093682ade20a9a6adbad99c5b3