Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 35 inbound Pith citation observations for arXiv:2303.03926.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:33:50.926162Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T00:04:22.393626Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c31f0988-d576-4e59-9964-a9dc3a1e4237 · inbound
The Rise and Potential of Large Language Model Based Agents: A Survey Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 272
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 979bebc8-0d3e-45fd-8ca7-22bb17eac15d · inbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8685f695-97b5-4632-b1b4-a12fb11e367b · inbound
F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 155
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 240147d3-4c80-4585-b45b-8d478dd37758 · inbound
Vox-Profile: A Speech Foundation Model Benchmark for Characterizing Diverse Speaker and Speech Traits Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64fe4a74-6347-4730-9265-6015099d6104 · inbound
Accelerating Autoregressive Speech Synthesis Inference With Speech Speculative Decoding Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e1872e8-7f8e-468d-bf75-a5acd1a15bc1 · inbound
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fab4a743-7c6c-42bb-a1ab-0ef3ae3e293a · inbound
Zero-Shot Streaming Text to Speech Synthesis with Transducer and Auto-Regressive Modeling Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f148466c-20d9-4340-a9ed-b9aedfd7a3f6 · inbound
Voice Adaptation for Swiss German Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ab4c217-426d-4035-a07d-ab117c4c1ffc · inbound
Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a7b921a-32ce-4f59-9765-7ed4dbf538f6 · inbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09225bfb-1e78-4efb-bf6e-b9e4b684cb17 · inbound
Kinship in Speech: Leveraging Linguistic Relatedness for Zero-Shot TTS in Indian Languages Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e58d9ddf-40d0-4d10-893d-ff661ecaf7eb · inbound
OpusLM: A Family of Open Unified Speech Language Models Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b87f9175-72f9-4052-ba6c-4c17668cba53 · inbound
Differentiable Reward Optimization for LLM based TTS system Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78b9d3b1-68ef-4d18-ae2e-3aabd768eb0f · inbound
SecureSpeech: Prompt-based Speaker and Content Protection Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4517a79-8173-4f65-8551-ffcd886f9f98 · inbound
Next Tokens Denoising for Speech Synthesis Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64a33c59-d966-4a47-b022-18342097d1b9 · inbound
XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9c9843c-e0b0-4d37-9727-f0540ccd7d32 · inbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dc06219-a9e9-4be9-b4e1-e28e44f6aed5 · inbound
DiFlow-TTS: Compact and Low-Latency Zero-Shot Text-to-Speech with Discrete Flow Matching Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c93a56ef-7037-49ee-95c9-33e0a5070a7b · inbound
CoMelSinger: Discrete Token-Based Zero-Shot Singing Synthesis With Structured Melody Control and Guidance Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 488ae25c-b8e7-4ae2-b441-d6cc262f33b1 · inbound
OmniCustom: Sync Audio-Video Customization Via Joint Audio-Video Generation Model Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84b0e2ae-172b-42a0-a79f-bcfdf1f1c229 · inbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a5fd5a40-14db-4ffa-b6e1-e1094da1f026 · inbound
StreamMark: A Deep Learning-Based Semi-Fragile Audio Watermarking for Proactive Deepfake Detection Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a70b231-b3aa-409c-87f2-0ca65bf922f5 · inbound
MoVE: Translating Laughter and Tears via Mixture of Vocalization Experts in Speech-to-Speech Translation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22772b80-26cd-4144-8f72-ad74a903bf89 · inbound
Human-1 by Josh Talks: A Full-Duplex Conversational Modeling Framework in Hindi using Real-World Conversations Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dddbbbb0-4ae5-4e67-9b5d-be997bff803b · inbound
X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8abf31f5-c8c4-4ebf-ba4c-3459173fd0dd · inbound
X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 11430ff3-1019-4358-86c2-e879e4a3d3d3 · inbound
Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f3ef0fae-32a5-48a0-bb78-c852c98aca91 · inbound
WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03a2684a-eeb9-4fbf-8e07-d08c633d1569 · inbound
Do speech foundation models perceive speaker similarity as humans do? Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b51ec42e-5a70-42bc-bcd1-f49e35a5c523 · inbound
UniVoice: A Unified Model for Speech and Singing Voice Generation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eff54b84-0495-4ced-a6c8-cb7fbd651d5d · inbound
CrossAccent-TTS: Cross-Lingual Accent-Intensity Controllable Text-to-Speech via Disentangled Speaker and Accent Representations Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23bea1f1-e3f1-4cb7-a49c-5a9264df1192 · inbound
VoiceTTA: Enhancing Zero-Shot Text-to-Speech via Reinforcement Learning-Based Test-Time Adaptation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03b696fa-7ee1-4897-96cd-1ece412df834 · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 235
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6fa88f4-b57b-4f87-ade2-356a59d72afa · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 235
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 487ed536-8ec9-4116-a1b7-b26693617478 · inbound
Component-Level Ensemble Fusion for Speech and Environmental Sound Deepfake Detection Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.