Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2305.16107.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:23:03.314763Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T00:04:22.442888Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c3e9db4e-2e37-4824-b626-ded906c5e4eb · inbound
DASB - Discrete Audio and Speech Benchmark VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b77fe242-bbaf-4c88-81c2-4b7a7701dbcb · inbound
F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 146
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 80f70929-f25f-419f-8682-d847b21bd568 · inbound
Mechanisms of Multimodal Synchronization: Insights from Decoder-Based Video-Text-to-Speech Synthesis VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f65ed92-a07d-4eba-b903-d475e502c3ad · inbound
Audio Jailbreak: An Open Comprehensive Benchmark for Jailbreaking Large Audio-Language Models VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc3048c9-080b-4860-b527-1fcdc0f87fc5 · inbound
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d051f26a-f4d5-4dc1-a1a0-71add2d27203 · inbound
Probing the Robustness Properties of Neural Speech Codecs VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4870119f-084f-4438-ab2b-8d42b446748a · inbound
NTPP: Generative Speech Language Modeling for Dual-Channel Spoken Dialogue via Next-Token-Pair Prediction VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82e7c989-1115-41fd-9d0b-10bca00a2b5c · inbound
Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f3ad701-dcea-4c3b-bad3-7a246672d405 · inbound
Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1716ce80-123e-484e-a4f3-65e6dab0be23 · inbound
The Alignment Curse: Modality Alignment Supercharges Audio Attacks via Text Transfer VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50646c69-958c-472e-8e22-427f1a6c492d · inbound
PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7080ea78-843c-43d7-8167-70a5797c4f18 · inbound
PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 982b8fe7-f2ce-4605-b23c-dcdaf6dfac20 · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 237
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 095db85d-7326-4c44-b4e9-7aebdf04307b · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
Reference 237
Source-reported events for the cited work
Unavailable: canonical work link unavailable.