Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 41 inbound Pith citation observations for arXiv:2310.00704.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:21:08.829562Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
14
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation e4360b2b-7a9d-4c88-ab5f-5dff3bc5fc37 · inbound
Moshi: a speech-text foundation model for real-time dialogue UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0c1be9db-808d-43b2-8d74-453dc411fda3 · inbound
Kimi-Audio Technical Report UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 48e7124c-05e4-43c9-bb0d-10ed4f8b5f88 · inbound
AudioTurbo: Fast Text-to-Audio Generation with Rectified Diffusion UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c87ec37f-a114-40b6-acff-57d356b935c6 · inbound
In-the-wild Audio Spatialization with Flexible Text-guided Localization UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eb02c9c-0049-4412-b20a-1d0f0b385f28 · inbound
Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93b38794-9dc5-4c3f-b4e9-219efef4dac9 · inbound
FusionAudio-1.2M: Towards Fine-grained Audio Captioning with Multimodal Contextual Fusion UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9596670-4fda-4713-bc55-68aeb0e506e2 · inbound
Foundation Model Empowered Synesthesia of Machines (SoM): AI-native Intelligent Multi-Modal Sensing-Communication Integration UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4572b879-d5a4-4980-9e0d-41350217fb08 · inbound
Scaling Laws of Motion Forecasting and Planning -- Technical Report UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 981006d2-d1a6-4647-8332-b80cd2f36eda · inbound
Exploring State-Space-Model based Language Model in Music Generation UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72949c76-f247-4187-949e-bf301ca95cc7 · inbound
SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d682d2a4-1638-47e1-b3b7-c04993354b10 · inbound
Generative Audio Language Modeling with Continuous-valued Tokens and Masked Next-Token Prediction UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6e79082-fc53-44c9-8f1e-6ec750618d11 · inbound
DualDub: Video-to-Soundtrack Generation via Joint Speech and Background Audio Synthesis UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 118c0019-ded0-417d-97e5-48b2b48b530c · inbound
Small Data Explainer -- The impact of small data methods in everyday life UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 506e6954-1a24-43c5-aaa2-a2c3c992b00e · inbound
Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff6b37b-4359-4374-8562-1d9f47b94cca · inbound
BoSS: Beyond-Semantic Speech UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecd473e6-b5a8-42f0-9157-10e01aae665a · inbound
Your Spending Needs Attention: Modeling Financial Habits with Transformers UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 404acb60-0c1c-40f7-93ad-3c6b2d3cc334 · inbound
Ego-centric Predictive Model Conditioned on Hand Trajectories UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6197b20-9dbe-4590-945c-9a1418149def · inbound
Enhancing Speech Large Language Models through Reinforced Behavior Alignment UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6ccd059-323c-4831-b26c-401f61f95b9c · inbound
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c4c84d1-c061-4e10-a641-3c466918b1cd · inbound
Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cefc0374-dafb-439b-9a9f-a548839f2bdf · inbound
Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b07c37f9-d40f-46f2-91cb-41d3251b0bbc · inbound
Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fb4263e-3399-4e4f-b0ee-b933c3288000 · inbound
iMontage: Unified, Versatile, Highly Dynamic Many-to-many Image Generation UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6f8bd97-d23e-4c84-ae04-9a6fff638289 · inbound
Qwen3-TTS Technical Report UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bbcbef27-93d7-41b0-986d-1c53182bed16 · inbound
Edit Content, Preserve Acoustics: Imperceptible Text-Based Speech Editing via Self-Consistency Rewards UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71031306-d505-4e03-98b8-542f9d4370e0 · inbound
Simultaneous Speech-to-Speech Translation Without Aligned Data UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f88f706-4546-4784-8303-9f521e2f6234 · inbound
Fine-grained Soundscape Control for Augmented Hearing UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f131168-718f-4581-ae3f-3eab37b2d5eb · inbound
Remix the Timbre: Diffusion-Based Style Transfer Across Polyphonic Stems UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c0fb6750-12f8-47aa-b72c-659f602c332e · inbound
Modeling Music as a Time-Frequency Image: A 2D Tokenizer for Music Generation UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2101993-a06f-4051-818d-136ce8d283bc · inbound
Unified Synthesis of Compositional Speech and Sound from Free-Form Text Prompts UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 323b2215-b214-4cd6-a90a-ff2a2dc5a7c7 · inbound
UniVocal: Unified Speech-Singing Code-Switching Synthesis UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f22033f-a2cb-4491-8224-66d7c5fdcebc · inbound
Self-Guidance: Enhancing Neural Codecs via Decoder Manifold Alignment UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 134
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7746aa85-7f8c-453c-b45b-f0bc2f686713 · inbound
NAC: Neural Action Codec for Vision-Language-Action Models UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e05d13f6-79a7-4bf1-a267-e5cfad63baa1 · inbound
AudioCALM: Continuous Autoregressive Language Modeling for Universal Audio Generation UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22c60912-0aec-4dea-9cd9-2e2be7da006e · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 59d0817a-b1a7-4eff-8df7-4707f059cc02 · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aa98018-0654-4444-959f-a554c30ce503 · inbound
ReGen: Hierarchical Multi-Prompt Representation Generation for Efficient Waveform Diffusion Models UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba1e27fc-e26b-4c2d-8126-fa61f865c25f · inbound
Qwen-Audio-3.0-Gen-Preview Technical Report UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8563ad7-2fc0-47f6-abc6-efa67a33c4ee · inbound
Qwen-Audio-3.0-Gen-Preview Technical Report UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6a686b1-0ac0-4803-9929-bcd8db3a2396 · inbound
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcd2f08e-ee7f-4ed1-b3c2-e482f20237cf · inbound
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.