Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T19:15:52.190529Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2502.05471.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T19:15:52.190529Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:58:58.610208Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T21:00:08.120813Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5d84f3fd-9b4c-4204-b572-9064d89a5838 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Autovc: Zero-shot voice style transfer with only autoencoder loss,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 93a89f50-f21d-4408-b5ae-cf42c10a674c · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model V oicemixer: Adver- sarial voice style mixup,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d8af58c6-fcea-4d97-8b41-58ab72fbedfd · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model One-shot Voice Conversion by Separating Speaker and Content Representations with Instance Normalization
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43547107-ed32-4a3a-b534-064397af18dc · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Again-vc: A one- shot voice conversion using activation guidance and adaptive instance normalization,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cb787d0-464c-4fcd-8407-52687a4eabd1 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Unsupervised speech decomposition via triple information bottleneck,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b245516c-ce93-4249-bb02-3e955d060569 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Neural analysis and synthesis: Reconstructing speech from self-supervised representa- tions,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a2649297-bf2f-4482-9e64-dad6d626e485 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Expressive-vc: Highly expressive voice conversion with attention fusion of bottleneck and perturbation features,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6441b8f8-fa2c-4e16-b2ed-17aafe308749 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Contentvec: An improved self-supervised speech representation by disentangling speakers,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3a85f7a5-9e41-456a-ba72-9a176f078019 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model UnitSpeech: Speaker-adaptive Speech Synthesis with Untranscribed Data
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38b95a36-cb90-4933-9d2a-9d3ab2c38828 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model A comparison of discrete and soft speech units for improved voice conversion,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1523f89e-a3ba-433c-910f-ad5e671b769b · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model OpenSR: Open-Modality Speech Recognition via Maintaining Multi-Modality Alignment
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 87f5e7dc-3889-406c-a99a-02e538b44f82 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Diff-hiervc: Diffusion-based hier- archical voice conversion with robust pitch generation and masked prior for zero-shot speaker adaptation,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 82359265-8bcb-441c-92b6-2f423f7bf549 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model TransFace: Unit-Based Audio-Visual Speech Synthesizer for Talking Head Translation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6aa15f87-6d3e-42c3-835e-378f81d778c7 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Hubert: Self-supervised speech representation learning by masked prediction of hidden units,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7aaf16de-cdf7-47dd-9d8b-0e43c321bd87 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad0c4ef2-79d1-43d7-8272-972afc4e003f · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Dgc-vector: A new speaker embedding for zero-shot voice conversion,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d57c9866-9809-4f2d-9c37-af4da096c778 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Zero-shot multi-speaker text-to-speech with state-of-the- art neural speaker embeddings,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9f88f6c7-6892-4eeb-8dbd-d307c5e12eac · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Sef-vc: Speaker embedding free zero-shot voice conversion with cross attention,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fda4c95-d68c-4d21-82d7-5bafa7c93a60 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Refxvc: Cross-lingual voice conversion with enhanced reference leveraging,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d393707e-dfb3-4b75-a6b8-12d9dc40ff9a · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Zero-shot voice conversion via self-supervised prosody representation learning,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e4cb2835-65ba-4e4a-839b-281ff92ae0f6 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Promptvc: Flexible stylistic voice conversion in latent space driven by natural language prompts,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e194f579-f819-468f-bf1b-ed4792528fbc · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model FluentSpeech: Stutter-Oriented Automatic Speech Editing with Context-Aware Diffusion Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93fd6717-fcae-4e11-8b67-9220276b82f5 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6340e8b2-c522-4926-9e68-620a185f2abc · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Speech Resynthesis from Discrete Disentangled Self-Supervised Representations
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38995372-dcbe-4898-9df2-0961d064be26 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc9bc02f-b76f-46c1-8358-ec84892ab68e · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Ace: A generative cross-modal re- trieval framework with coarse-to-fine semantic modeling,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaa02832-2131-4664-832e-6c0440586611 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 287fa4db-112d-461a-bf0b-212ed2e208de · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75f19251-7957-4dfc-8a97-d9e9c0bf4ed3 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Neural ordinary differential equations,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3d6a598e-5eb3-4808-a4fa-9a2d80d7b681 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Flow Matching for Generative Modeling
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39624ee8-c5e4-474c-bd57-5436f2d96644 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9459da-28bb-4a7e-9ac9-3aa110f5e83f · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0efcc18-4f6e-4ecb-ae43-e1f71c48e62b · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Librispeech: an asr corpus based on public domain audio books,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce3c3db8-3abc-4b9d-b97d-7b59327a41d8 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Seen and unseen emotional style transfer for voice conversion with a new emotional speech dataset,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a5f6dbd-87cb-464d-b939-b4990f0428ed · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Matcha-tts: A fast tts architecture with conditional flow matching,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 143e4c8e-f340-4307-b93c-822ef83fa9be · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model emotion2vec: Self-Supervised Pre-Training for Speech Emotion Representation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c2ebe9d-6be7-41f6-a18d-7d3b966b2fb1 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model F0- consistent many-to-many non-parallel voice conversion via conditional autoencoder,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7e8479f-ed8a-4ef3-ab44-15efc7b8ad70 · outbound
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model Text-Free Prosody-Aware Generative Spoken Language Modeling
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74f6896e-0381-4b69-8a07-941b1548c3bc · inbound
EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69fa6331-1fd0-4995-8696-417777240f33 · inbound
Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8432b24e-8c02-4f75-883e-c008952b50d2 · inbound
SpeechAccentLLM: A Unified Framework for Foreign Accent Conversion and Text to Speech Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.