Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:19:03.875806Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 2 inbound Pith citation observations for arXiv:2506.22023.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:19:03.875806Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-27T01:05:20.432379Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:58:57.478873Z
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a5e68496-bb31-450b-84ca-0952e784f0c2 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy A Survey of Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb1ccaf6-217f-4166-bf0c-6b91003ee642 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Codec-SUPERB: An In-Depth Analysis of Sound Codec Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd8c8cb0-03ac-476c-8e5b-b72ad47d999e · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Recent advances in discrete speech tokens: A review,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ef43034-e09c-47b0-b03f-f0e7494a5942 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy On The Landscape of Spoken Language Models: A Comprehensive Survey
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 170c9bb8-d83e-4618-bde3-c75ad123e85e · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Neural Discrete Representation Learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d465b53-a010-46e1-83f7-f162dcbfc1dd · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Fast and high-quality auto-regressive speech synthesis via speculative decoding,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7cfcf5d-e68c-4595-8ce2-625a9860e24b · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Accelerating Codec-based Speech Synthesis with Multi-Token Prediction and Speculative Decoding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f734fbc-4e28-473c-b2b5-189bfff8215b · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy VocalNet: Speech LLM with Multi-Token Prediction for Faster and High-Quality Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b95c2a2-7701-426e-9b56-1e2a926f9b7d · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac48b1a8-3cfd-458a-8666-e40d88a486a5 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0eaed673-25f5-44ed-9d4a-9084025b75a2 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10172a57-2cf3-4b8f-a1e9-81406f67e2dd · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy LibriTTS: A Corpus Derived from LibriSpeech for Text-to- Speech,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b04355ef-2ce0-4264-8d3c-142da83fbb8a · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f64dcb36-58b7-422b-83b9-42bbada1cba2 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cca8f652-0bf6-4cff-9abb-116298515508 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbe6412f-31f3-4fdc-b33f-731c6b255e13 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy V oiceCraft: Zero-Shot Speech Editing and Text-to-Speech in the Wild,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3fff0af9-2b2b-4938-aa93-41cff71acb69 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45eb768c-4d27-4724-8018-1a3d215099e3 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f544de7-5d73-42a3-aa1b-fde17836f6c3 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy ELLA-V: Stable Neural Codec Language Modelling with Alignment-Guided Sequence Reordering,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da940102-74c5-4465-8666-46c1934a7c50 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy VALL-E R: Robust and Efficient Zero-Shot Text-to-Speech Synthesis via Monotonic Alignment
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b915efd-50a3-4683-9d04-669e86c5673f · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy V ALL-T: Decoder-only generative transducer for robust and decoding-controllable text-to-speech,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b967284-e98e-45fc-b7af-9a6177e2998a · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy RALL-E: Robust Codec Language Modeling with Chain-of-Thought Prompting for Text-to-Speech Synthesis
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c9ca762-3128-47f8-95e8-b181b5323ef8 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy SNAC: Multi-Scale Neural Audio Codec
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ff7cc4e-5338-474b-a2b5-c427656fce26 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Speaking from Coarse to Fine: Improving Neural Codec Language Model via Multi-Scale Speech Coding and Generation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 453300fe-b9e7-4766-bb5e-e29390984124 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy UniAudio 1.5: Large Language Model-Driven Audio Codec is A Few-Shot Audio Task Learner,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 623308b0-5410-4c93-921d-76604959d8cc · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy SpeechTokenizer: Unified Speech Tokenizer for Speech Language Models,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 541ac633-3efc-459c-82c6-140d69c2f425 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Moshi: a speech-text foundation model for real-time dialogue
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f31aca70-c936-420a-b30f-4a1e6c6ec694 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy AudioLM: A Language Modeling Approach to Audio Generation,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87d7d26a-7d45-4b9e-a157-954651f6f045 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Speak, Read and Prompt: High-Fidelity Text-to- Speech with Minimal Supervision,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 212b35dc-657a-48f6-ae02-a3db385f2a66 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 021b628d-e59d-41e5-9d24-abc2e64e2af3 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36d5de7f-a8ec-4f5a-bd80-308858dd3cb8 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Better & Faster Large Language Models via Multi-token Prediction
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01db0498-5f06-4f96-ac2c-35d8cbbe18d4 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy DeepSeek-V3 Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 034f4c16-c96a-4c01-8782-c94c93b10491 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Training language models to follow instructions with human feedback,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1f83a33-32f4-4d46-ad4d-407412ceb971 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Proximal Policy Optimization Algorithms
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b2ec461-051b-48dd-a584-65794bf75d01 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Direct preference optimization: Your language model is secretly a reward model,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed5e40cf-f30d-403d-b151-12f36ac6cea7 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b38bdcb0-3913-450b-8c11-8b7e27459dd7 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Align-SLM: Textless Spoken Language Models with Reinforcement Learning from AI Feedback
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76cb5b0b-a16d-4a52-87f3-b884024da7cf · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Robust Zero-Shot Text-to-Speech Synthesis with Reverse Inference Optimization
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff5959b-423b-4aed-a646-6c0e663d169a · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Enhancing Zero-shot Text-to-Speech Synthesis with Human Feedback
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 576192fd-545d-423f-a2fd-742d69e6552d · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy SpeechAlign: Aligning speech generation to human preferences,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71a4a6ff-5a9b-4326-9b1f-18cdfee7b83c · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Fine-grained preference optimization improves zero-shot text-to-speech,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a995078-6c79-4198-9ada-73e1eedc2b32 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy VALL-E 2: Neural Codec Language Models are Human Parity Zero-Shot Text to Speech Synthesizers
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9935a8d-bf66-4312-a1fc-ca1e4ab6f111 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy UniCATS: A Unified Context-Aware Text-to-Speech Framework with Contextual VQ-Diffusion and V ocoding,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e51de082-c27f-4ccf-9cef-17c1e6e69518 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy LSCodec: Low-Bitrate and Speaker-Decoupled Discrete Speech Codec
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65daf919-b05e-4438-8033-4f3a98cb4407 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8c83b8d-7981-4eae-a62d-7fefc3f3a776 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Sigmoid-weighted linear units for neural network function approximation in reinforcement learning,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2d0d98b-d896-40cd-99c4-b34e37f99f92 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy UTMOS: UTokyo- SaruLab System for V oiceMOS Challenge 2022,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 75ad788d-482b-47b3-9388-04dc6c7cf032 · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Robust Speech Recognition via Large-Scale Weak Supervision,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5835c07c-950d-4118-957f-594a85fa128b · outbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Zipformer: A faster and better encoder for automatic speech recognition
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03bae24a-489b-4704-8727-8f4f7b6df872 · inbound
From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5d064c5-efdc-4ec2-aab7-be973da6cf34 · inbound
ASTRA: A Scalable Next-Generation ATCO Training Simulator with Autonomous Simpilots Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.