Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:11:55.394112Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 2 inbound Pith citation observations for arXiv:2507.18897.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:11:55.394112Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T05:20:49.952030Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
25 of 25 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 462a63bc-6a3d-4e3e-bf9d-c9d5f4b96dc3 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23e1e775-e1b9-4f7d-a881-f28be003e22b · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Overview of the evs codec architecture
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fe1a2dd0-d4b6-4767-882d-e715fcaf1d60 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Textless Speech Emotion Conversion using Discrete and Decomposed Representations
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20bdf8b8-6bc0-486e-8ee6-fbf87c9fa810 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Multi-scale sub- band constant-q transform discriminator for high-fidelity vocoder
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation aea01c95-f0f8-475a-b9e9-fd2a39d1bcb2 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Emilia: An extensive, multi- lingual, and diverse speech dataset for large-scale speech generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69485992-1638-4842-b044-579ad86e708e · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb1a8488-8f32-4b5f-9aca-b31bc52845b5 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling BigVGAN: A Universal Neural Vocoder with Large-Scale Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edbdc2fb-8313-446b-aea5-3ed0bb6e2846 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Single-Codec: Single-Codebook Speech Codec towards High-Performance Speech Generation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48a87d0d-3442-44be-9bd9-331e8abd7710 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Finite Scalar Quantization: VQ-VAE Made Simple
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37aa6348-e795-4132-9b4d-342faa23e802 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Enhanced Direct Speech-to-Speech Translation Using Self-supervised Pre-training and Data Augmentation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9779287d-7d2f-4980-a85e-2d13418a2679 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc83373a-0ebe-4efa-9419-006e75ff1fbc · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Neural Machine Translation of Rare Words with Subword Units
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eb0fb14-fa1f-4604-a4f8-8687aee6342d · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 410dfe13-3645-4307-afc3-b8fc8a6f6117 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling H., Hendriks, R
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bfaf7900-0d04-4e9e-a88a-0bfc24356ab2 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecaa0243-654e-4c06-8a15-9b880b43ace4 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Addressing representa- tion collapse in vector quantized models with one linear layer
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d92e18f-e846-483e-bbbd-016068783d3e · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling SEANet: A Multi-modal Speech Enhancement Network
Reference 2010
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3cd9dc7-a8cc-4d21-84b1-b77f000f398c · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f8e55ff-669e-4306-99a9-5fc2d3ffc1f0 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 784dd7a5-97e1-4b6f-8cf6-d14127b47a93 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Virtual Class Enhanced Discriminative Embedding Learning
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 623e8020-74cd-4149-ac46-cf8d4d9fafba · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d85cbd5-7560-4c54-8f46-8d607f89fb94 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Restructuring Vector Quantization with the Rotation Trick
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f94f1c66-4099-4ddc-a52c-ce56e9e9c916 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Moshi: a speech-text foundation model for real-time dialogue
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a16ce14-c7d5-4a2c-bd3a-3876edc30654 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a54e8fc2-36dc-4a05-aa11-07d2d9aa33c8 · outbound
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling Vision Transformers Need Registers
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec4ecc8f-4c80-40da-9c5e-df10243578cd · inbound
Listen, Pause, and Reason: Toward Perception-Grounded Hybrid Reasoning for Audio Understanding HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4fd98e9a-bfcf-4e56-bc5f-8fea8212fb17 · inbound
CleanCodec: Efficient and Robust Speech Tokenization via Perceptually Guided Encoding HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.