Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:17:06.772619Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 2 inbound Pith citation observations for arXiv:2508.02038.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:17:06.772619Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T12:54:20.815371Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T01:06:23.916814Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d657cd9c-ea7a-489f-b957-772d85182335 · outbound
Marco-Voice Technical Report Barakat, O
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a6328654-ac4e-4765-a3e6-51717d90e619 · outbound
Marco-Voice Technical Report EmoSpeech: Guiding FastSpeech2 Towards Emotional Text to Speech
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30badf42-c085-493f-ba47-67cea15fd789 · outbound
Marco-Voice Technical Report Cross-speaker Emotion Transfer Based On Prosody Compensation for End-to-End Speech Synthesis
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5128c4bc-6773-4381-bdbf-52782b84470e · outbound
Marco-Voice Technical Report EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39a25de5-0fe5-4e27-b05d-6ca480cb2e41 · outbound
Marco-Voice Technical Report DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3dd2c9a9-4ed4-4bde-be05-e8b4c8dbfb64 · outbound
Marco-Voice Technical Report NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6834a970-203c-428b-8fbb-b7a01beff663 · outbound
Marco-Voice Technical Report A Survey on Neural Speech Synthesis
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9505d7d4-4fef-46d4-9a41-e7b1e62bcb5b · outbound
Marco-Voice Technical Report Improving and generalizing flow-based generative models with minibatch optimal transport
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 605f7043-c24c-42d9-82e5-de471472d079 · outbound
Marco-Voice Technical Report A Survey on Audio Diffusion Models: Text To Speech Synthesis and Enhancement in Generative AI
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0776a0c-c7fe-4d93-9961-eff89e5c652a · outbound
Marco-Voice Technical Report Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 88e9e5bb-5305-4156-89a3-d955a6123bc4 · outbound
Marco-Voice Technical Report Unresolved cited work
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 867aa1ca-1fb9-4924-9e32-ac2e1e0b9e10 · outbound
Marco-Voice Technical Report Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38364a48-0d8c-4fa9-956d-10e6cce7a254 · outbound
Marco-Voice Technical Report Spontaneous Style Text-to-Speech Synthesis with Controllable Spontaneous Behaviors Based on Language Models
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7272df6f-6938-436b-b387-22b64fdb9620 · outbound
Marco-Voice Technical Report CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 171d1df3-2360-4aa0-bc5b-a4d07157c2bd · outbound
Marco-Voice Technical Report ERes2NetV2: Boosting Short-Duration Speaker Verification Performance with Computational Efficiency
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c5e8dea3-6b19-48c0-a1b7-2b0eca1b5b19 · inbound
StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs Marco-Voice Technical Report
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ae97962b-e266-4109-8d59-ed4aa6f9e464 · inbound
SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing Marco-Voice Technical Report
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.