Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:40:10.557940Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2509.01391.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:40:10.557940Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5ff54f5e-682b-4d85-bd31-17cf031384a9 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model A Brief Overview of Unsupervised Neural Speech Representation Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 90f5f69d-626c-4cd5-ae28-15d4a175a960 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Natural TTS synthesis by conditioning Wavenet on mel-spectrogram predictions,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8e44e2db-0d4c-446a-8314-cd30ad8b20d0 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9172d4dd-ad1a-421f-b1c9-3208c80ac016 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model On generative spoken language modeling from raw audio,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e34ba3e0-1027-4ac8-8070-ab1221fa4917 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Wav2vec 2.0: A framework for self-supervised learning of speech representations,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c6abd5db-d1d3-4674-9079-3c8f01a02db8 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model HuBERT: Self- supervised speech representation learning by masked prediction of hidden units,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5e76a905-5d1d-4d2e-a1bb-11c2f2f79f8e · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model A Unified Accent Estimation Method Based on Multi-Task Learn- ing for Japanese Text-to-Speech,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c4c91880-fe62-47fb-9d30-c05813665721 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Reazonspeech: A free and massive corpus for Japanese ASR,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation daf8e8ad-0319-407d-99de-c4f775461830 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Exploring the limits of transfer learning with a unified text-to- text transformer,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4326f94e-a02e-415e-a5c6-8d8e5a9fc790 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7237d5cf-fe98-43df-a014-99330b8e6193 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model JVS corpus: free Japanese multi-speaker voice corpus
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02a76c6d-7964-40a0-8132-93a3e365dd55 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Audio- book speech synthesis conditioned by cross-sentence context-aware word embeddings,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9d4ba3ca-6227-4aa6-8e29-535cc43dc83b · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model J-MAC: Japanese multi-speaker audiobook corpus for speech synthesis,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f41b6dad-0f1d-4279-8df4-09aa58a892d9 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model JSSS: free Japanese speech corpus for summarization and simplification
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d808ecc-efea-487b-9f18-88eb319e06ef · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model T5g2p: Us- ing text-to-text transfer transformer for grapheme-to- phoneme conversion,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5af588d6-e2ec-44fe-819d-ace2c0e3a93e · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model SpeechT5: Unified- modal encoder-decoder pre-training for spoken language processing,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 36323435-2a5d-4930-81db-cf68a5d70b52 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Neural ma- chine translation for multilingual grapheme-to-phoneme conversion,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 13135d83-6c53-4f40-81dc-c563038ebeb5 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model One model to pronounce them all: Multilingual grapheme- to-phoneme conversion with a transformer ensemble,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d50b1644-9ac2-43ee-bb72-12798424b365 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Byt5 model for mas- sively multilingual grapheme-to-phoneme conversion,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 369ce2af-1ef8-40ed-84bf-32853dd86534 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model ContentVec: An improved self-supervised speech representation by dis- entangling speakers,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1257031f-77b7-42fc-a4b6-f15c4431988c · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model FastSpeech 2: Fast and High-Quality End-to-End Text to Speech
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7c8b8da-ac01-4958-b309-6a285538d60a · outbound
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e515bfbf-108a-4b89-929a-49b24a59a620 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model UTMOS: UTokyo-SaruLab system for VoiceMOS challenge 2022,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2ae9a126-9e21-46af-a0d5-77bf7ccb7e17 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Speech quality assessment with W ARP-Q: From sim- ilarity to subsequence dynamic time warp cost,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 90e04134-54e2-4c17-8272-80191bbcc018 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model Per- ceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 07250b96-e2a1-4327-9ce1-90001fc03b31 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model mT5: A mas- sively multilingual pre-trained text-to-text transformer,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d9cf1ded-3c45-4a9f-957d-e7bd01a8cb69 · outbound
MixedG2P-T5: G2P-free Speech Synthesis for Mixed-script texts using Speech Self-Supervised Learning and Language Model ByT5: Towards a token-free future with pre-trained byte-to-byte models,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
No inbound Pith citation observations are available.