Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:02:45.995503Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2505.11200.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:02:45.995503Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-08T12:03:11.243693Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T19:21:09.992559Z
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d12e0eee-0575-4604-92f3-acec6972d62a · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese The kendall rank correlation coefficient
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 78e42f87-1c18-4a06-bbc6-b9bf84f4f46d · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49ab2a14-7b56-4a04-961c-c6bec3e765b2 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese The t05 system for the voicemos challenge 2024: Transfer learning from deep image classifier to naturalness mos prediction of high-quality synthetic speech
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 27821350-9168-49bc-855e-bdb93b82c704 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Generalized linear mixed models: a practical guide for ecology and evolution
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2feb102f-9eb0-4a89-9a40-03be0124b9db · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Why we should report the details in subjective evaluation of tts more rigorously
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 34a5caeb-448c-4616-b27d-41c5bec6b7c2 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Qwen2-Audio Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54564007-9d02-4775-8abd-164db0a99811 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd94a023-0738-4b32-85ff-82892eab48c6 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Disambiguation of Chinese Polyphones in an End-to-End Framework with Semantic Features Extracted by Pre-trained BERT
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c987edd5-03fb-4ff8-bef5-4dd278e08814 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 939b2ce8-a515-4760-b107-dd68c9163cea · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Assessing the impact of contextual framing on subjective tts quality
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 075d394f-af28-4007-8479-c4a61a262681 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese The turing test: the first 50 years
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0bcd6c3c-53d4-4c3a-ae1a-db140529cb68 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Analysis of speaker similarity in the statistical speech synthesis systems using a hybrid approach
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3f9bdd01-ed8a-49c2-83cc-865a0096be17 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 838d049d-4211-44f4-8e20-80ca2dd798ed · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Lora: Low-rank adaptation of large language models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0873eeb2-a2ab-40a1-97d6-4a8611358ce0 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2c57a22-089a-4124-a412-99066bacda59 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Mm algorithms for generalized bradley-terry models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc5ef76a-f15e-4309-ab88-086bec1649de · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese GPT-4o System Card
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c87f2b6c-f133-4c6f-9ddb-07c0a1ee6fa1 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Method for the Subjective Assessment of Intermediate Quality Level of Audio Systems, 2015
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 71107992-3aec-49a1-a264-a01315bd3f0b · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Subjective evaluation of speech quality with a crowd- sourcing approach, 2018
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8918ce2d-48dc-44f8-83a0-fb5157157b10 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Compact Neural TTS Voices for Accessibility
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4fe8b601-769d-4554-adff-150867f3c357 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Stuck in the mos pit: A critical analysis of mos test methodology in tts evaluation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 78dc6c6c-af68-4195-ae55-065e6914a91e · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Issues in chinese prosody: conceptual foundations of a linguistically-motivated text-to-speech system for mandarin
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1ddbb4e4-8785-4c2b-9de7-9bd1c2c331a4 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese The limits of the mean opinion score for speech synthesis evaluation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 25e06993-00be-4795-9ff1-7e7478a64e9e · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee85daee-7f91-4448-bb93-d3020af5c143 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Hyper-realistic, multi-emotion generative speech model speech-01
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 23f73e61-8485-4b1d-b7a4-b63c8ebd9b16 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Speech Quality Assessment in Crowdsourcing: Comparison Category Rating Method
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 653dc60e-7403-48af-a90e-c356fc618f18 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese The blizzard challenge
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5ecee893-2d36-4eca-a73a-a04941c003c4 · outbound
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a8de6606-9795-4a78-bbc8-7041e0f01c01 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ec24e69-2d0f-447d-9579-242ec86fb6cf · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Mean opinion score (mos) revisited: methods and applications, limitations and alternatives
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2d921818-f727-435a-b522-005c737f94ca · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eca95645-6b86-4f9c-8307-7e66d4703efa · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed67bc23-3e89-4bb0-bb63-f9fcfed3ee86 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Contextual interactive evaluation of tts models in dialogue systems
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 320ec6cd-47ca-4755-b7f4-245507619faa · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfbe0640-77c6-4aff-81a3-0eb3fbfc9e74 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Bilingual and code- switching tts enhanced with denoising diffusion model and gan
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7d5914c6-f227-43b5-876d-6e43de4d5959 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Talk2care: An llm-based voice assistant for communication between healthcare providers and older adults
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e0df59f-fc67-4124-91d2-ab4338032a0c · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Dialog modeling in audiobook synthesis
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 730a5761-7d6e-48b7-b7f6-fc5a6f450973 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese orders of magnitude larger
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cdbc3f11-774e-4e02-b821-670445f8fc91 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Pure machine voice
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a6cc27a3-e201-445b-8c9d-6e51f2a837f3 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese The i m i t a t i o n of human speech is too forced
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 64a8d3f5-6982-4628-b4c9-dae1e669ed74 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese O b v i o u s l y a machine tone - doesn ’ t sound like a real person
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation da151159-e479-4e0f-999d-208f638af4c1 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Sounds like a late - night radio host
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ac974039-4221-4c92-afeb-5000f6d26a52 · outbound
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cc2d41db-d096-4f5e-b35b-038e7cb39e49 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 69e6cc48-396f-48e5-ae18-3bec302d2417 · outbound
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 730e009a-eede-46dd-b78a-3a67896de368 · outbound
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 45f9dab7-753f-4a68-a415-96c670f0e988 · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese doi: 10.21437/Blizzard.2023-1
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1aed6274-3448-4c8d-afe2-64eb5efbe64b · outbound
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese URL https://www.sciencedirect
Reference 2308
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1e43662-3b5f-4805-b13a-e882f93a10a2 · inbound
TTS-PRISM: A Perceptual Reasoning and Interpretable Speech Model for Fine-Grained Diagnosis Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.