Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T15:10:46.729901Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 2 inbound Pith citation observations for arXiv:2501.14477.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T15:10:46.729901Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:37:08.860549Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T14:22:07.662926Z
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e64515a8-7a55-4d01-a669-1be15a21e8d1 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Supervised speech separation based on deep learning: An overview,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e8c4950-15d6-425b-a325-e7a2318b43bb · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR SpeakerBeam: Speaker aware neural network for target speaker extraction in speech mixtures,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fcfe31bd-1c2b-4892-b1df-dae02062b0af · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Improving speaker discrimination of target speech extraction with time-domain speakerbeam,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 93bdc0e6-9656-4e04-ae97-cd9bbbb1a8f9 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR SpEx: Multi-scale time domain speaker extraction network,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4c717d8-3fbb-445e-8cac-eea030e9a485 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Target speech extraction with conditional diffusion model,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a8da5e04-21df-499d-88c1-a9a7da46c977 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Generation- based target speech extraction with speech discretization and vocoder,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 72fe91f6-2256-4f34-859b-83c10e4fe051 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR TSELM: Target Speaker Extraction using Discrete Tokens and Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d37f81d3-ac7a-4cde-8dce-ef0c2d00def9 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR X-Sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d282659c-dd21-4b2d-8c79-85235707f1db · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Personalized speech enhancement combining band-split rnn and speaker attentive module,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 13178b6d-e481-4a2e-9227-e84482ef5244 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Multi-Level Speaker Representation for Target Speaker Extraction
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc4661f7-00c1-45cf-b940-995b05860785 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Hierarchical speaker representation for target speaker extraction,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f9ef6128-fec9-402f-b9c4-c32700df0e34 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Extending Whisper with prompt tuning to target-speaker ASR,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ed4a4fa8-2d2c-479b-b71b-863bd58dde34 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Empowering Whisper as a joint multi-talker and target-talker speech recognition system,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e342ca0a-6445-478a-bb8e-179f59ca1356 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR SQ-Whisper: Speaker-querying based Whisper model for target-speaker ASR,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 17239752-4b59-4f8a-a1cf-0f039e3fa314 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Robust speech recognition via large-scale weak super- vision,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e9c76bbf-e0b0-45ce-906c-83d722439235 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR LoRA: Low-rank adaptation of large language models,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 96b622a0-1d44-4d2e-b856-4e7e6a118a3d · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Flow matching for generative modeling,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4c93146-3020-42a2-9c87-b2d2bfd1be34 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Matcha- TTS: A fast TTS architecture with conditional flow matching,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6fa45df7-e2a2-428e-ae49-fbae5c7f5d18 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd89a5ae-8e9b-4f10-b316-0089f7cd367d · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR LibriMix: An open-source dataset for generalizable speech separation,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e22286fa-806d-41bd-b8aa-2d4b69a430bd · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Deep clustering: Discriminative embeddings for segmentation and separation,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 25e09a36-2e1e-4397-a44f-14f7a1b8b9b7 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Attention is all you need,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d860a433-1951-4857-ab67-e1a204d7d755 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Drop the beat! Freestyler for Accompaniment Conditioned Rapping Voice Generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1f28baf5-cf7d-4094-997d-1e7b7545b1e9 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR StableVC: Style Controllable Zero-Shot Voice Conversion with Conditional Flow Matching
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0eb6be9-90a9-41bf-a34e-fc2692bc919e · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR HiFi-GAN: Generative adversarial net- works for efficient and high fidelity speech synthesis,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 23f16fff-4312-4732-a716-753ebda57677 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR LibriSpeech: An ASR corpus based on public domain audio books,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 082f1ea6-a0ad-41f0-b487-b0394f311f39 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Adapting self- supervised models to multi-talker speech recognition using speaker embeddings,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ea68ea39-dc31-43a3-b84c-dd07cc4a4ceb · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Single channel speech separation with constrained utterance level permutation invariant training using grid lstm,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b8afd31f-848a-4962-82aa-76c6d4097dc5 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR DNSMOS P.835: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1c594bd9-7a28-49c4-9b0c-fd684e0db7b3 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5720d9b5-9734-4c39-b064-5348efef89f8 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR CAM++: A fast and efficient network for speaker verification using context-aware masking,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d1452fe5-15e0-49d6-9c63-f139780b8ebf · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR SELM: Speech enhancement using discrete tokens and language mod- els,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ce9906db-1e5f-4ad5-8f8a-f117e3369987 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Decoupled weight decay regularization,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b19be8a5-f6a0-4511-a199-e0cdff379eb3 · outbound
Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR Music source separation with band-split rope transformer,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation df8b61db-204e-4fd2-b94b-825f583d3405 · inbound
FlowTSE: Target Speaker Extraction with Flow Matching Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25dbd4bf-1922-457e-abd1-dd3165775919 · inbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.