Pith. sign in

Paper Citation Record · LEDGER

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

As of 8 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 1 inbound Pith citation observation for arXiv:2509.09201.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09201 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:37:53.994131Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:37:50.680596Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved55
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 418f8ce1-a402-44c1-90ce-25eb16f76ea9 · outbound

This paper cites DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.680596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.680596Z digest=sha256:9320332881fc36fcba66dcf96470b4faefce771259548bbac1319a6a432e5571

Observation 51ebac8a-437c-4215-9974-fe6f2ad2e220 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.725994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.725994Z digest=sha256:23c5fe4c09f2a708d02eafb176157ed83adf7e669c48e475be7509fc53e485cc

Observation d8b17411-afa8-4976-9d98-db076111a33b · outbound

This paper cites Based on the differences in the physical generation mechanisms, it can be assumed that speech signal and background sound signal are mutually independent [31].

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Based on the differences in the physical generation mechanisms, it can be assumed that speech signal and background sound signal are mutually independent [31]

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.757211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.757211Z digest=sha256:910310d49c012c855c76148ea5347a213d3a8e497ba583993975a6f8fd545464

Observation d759351c-a47f-49b0-ae0f-79cb43032097 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-08-04T19:37:50.818276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.818276Z digest=sha256:eea1457d41ab2af137045e7087e0bd4a2423c4de23ef2ce4816e2047a9e812b0

Observation 72f25802-bf56-4d13-895e-b50624272a68 · outbound

This paper cites front-end separation + back-end processing.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners front-end separation + back-end processing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.870202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.870202Z digest=sha256:5f25ff3b9adfdf4be971ac3454dc2aeb28482859cbe4dd4e0778cf3aa4f000d1

Observation 767b99ed-9000-425b-bb18-2a081141b7e3 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 6

Resolution
malformed identifier
no resolver link, observed 2026-08-04T19:37:50.927592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.927592Z digest=sha256:666b856b8ad482ef8f0ec29606bc486f081d04e9b3e40d0f8ed159e1780f45f0

Observation d74fee62-8daa-43f8-866d-6477696a3bba · outbound

This paper cites The experimental results confirm that the representations are sufficiently disentangled, enabling controllable feature selection tailored to diverse downstream tasks.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The experimental results confirm that the representations are sufficiently disentangled, enabling controllable feature selection tailored to diverse downstream tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.015687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.015687Z digest=sha256:09c512721ad29c1968178e268f71ece6a480cdc3b988c2b3354ec09dedd435b8

Observation 3e72c3ac-f75b-48d2-aed9-0dd507bb38bf · outbound

This paper cites AudioPaLM: A Large Language Model That Can Speak and Listen.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners AudioPaLM: A Large Language Model That Can Speak and Listen

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.101392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.101392Z digest=sha256:0f396cf5f138bf9e941ff9c920a52b982670025326a3091fdc533b0e02c02183

Observation a000a9fd-0dd0-4ee8-9670-f95db715c479 · outbound

This paper cites Audit: Audio editing by follow- ing instructions with latent diffusion models,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Audit: Audio editing by follow- ing instructions with latent diffusion models,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.183385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.183385Z digest=sha256:8ea84938f5052f67035af807ab85fcc552ebe88329d8046fb74b3ade03d8fcad

Observation 0093045f-152c-4769-ac0d-64635a00c715 · outbound

This paper cites Superm2m: Supervised and mixture-to-mixture co-learning for speech enhancement and noise-robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Superm2m: Supervised and mixture-to-mixture co-learning for speech enhancement and noise-robust asr,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.260311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.260311Z digest=sha256:ab15f1e946a0d9475edbb0654d85aca8eec85f5352735eb6262f6814d6f2033f

Observation 8c819585-5800-4c7e-8115-d2922fb1edba · outbound

This paper cites Preserving background sound in noise-robust voice conversion via multi-task learning,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Preserving background sound in noise-robust voice conversion via multi-task learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.322124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.322124Z digest=sha256:328c7af02df7ddc019c53e61b752a1cd3a4517265c8f206a3644617844563a4c

Observation c2c04314-f074-4cb2-bc94-9e63231eff77 · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Supervised speech separation based on deep learning: An overview,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.405886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.405886Z digest=sha256:b44294ad863c48700e36a4191c333e4d094e113b6c7ea0753797b5dca646ce07

Observation 675636fe-b5ad-4936-ae5f-1fda4a9b4368 · outbound

This paper cites Complex spectral mapping for single-and multi-channel speech enhancement and robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Complex spectral mapping for single-and multi-channel speech enhancement and robust asr,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.476825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.476825Z digest=sha256:4373d4f2ee2fef1f9c06e2254c93f2abb1cccbeac094e100b8dcc3df9449c6a2

Observation 7636640a-1529-434f-8adf-e9212ee8de43 · outbound

This paper cites Speech enhancement with lstm recurrent neural networks and its application to noise- robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Speech enhancement with lstm recurrent neural networks and its application to noise- robust asr,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.519961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.519961Z digest=sha256:fdcdedced98b5f5fd1aeac357caa578cc8309bb9b03756d14bdb57a86dea77f5

Observation 0919e23d-e6ec-4f44-ba40-89530926a584 · outbound

This paper cites Deep neural networks for speech enhancement and speech recognition: A systematic review,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Deep neural networks for speech enhancement and speech recognition: A systematic review,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.571480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.571480Z digest=sha256:60183390680d6b9ef68e1b0716fe0546a6d75a74d3829d083c5ef0a9595ae4cb

Observation 8e8b811f-e6a9-42c5-adea-5d923a969839 · outbound

This paper cites Attention-based latent features for jointly trained end-to-end automatic speech recognition with modified speech enhancement,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Attention-based latent features for jointly trained end-to-end automatic speech recognition with modified speech enhancement,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.650758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.650758Z digest=sha256:cacd35fa6df7ff5dca09279f1d564a3ecf0d416bac45d989cbf8fd48762de6dd

Observation 0b3eaef0-d7cd-412e-89ea-c3cb7eabf06b · outbound

This paper cites Phonetic feature encoding in human superior temporal gyrus,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Phonetic feature encoding in human superior temporal gyrus,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.734805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.734805Z digest=sha256:18cfd18785d1e66ad41baa0d88020fefc23bf957931104c8642dddc2afba40fd

Observation 304a0721-9c83-40a0-acce-11fd71c31963 · outbound

This paper cites Representation of temporal sound features in the human auditory cortex,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Representation of temporal sound features in the human auditory cortex,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.791267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.791267Z digest=sha256:1185759b4ba9b2458d8f9031aa73f31eccb2305972c84ff850f6b9ca162d23a5

Observation 288879da-e5ce-4acf-9ef9-f8d101a3c88a · outbound

This paper cites High Fidelity Neural Audio Compression.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners High Fidelity Neural Audio Compression

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.875056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.875056Z digest=sha256:d21519905cff35cf5683baa6623563396b196960a3b74751267066c749e279a0

Observation a70cc88a-458a-4e35-bd7a-f1b01a6c237e · outbound

This paper cites High-fidelity au- dio compression with improved rvqgan,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners High-fidelity au- dio compression with improved rvqgan,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.951394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.951394Z digest=sha256:f989e3b99a4a33c5cbf9655baf3cff4da45d2465d26d3b77ba773d33bb275649

Observation 143f1e17-f16f-4f2d-b198-ec8a76fb5623 · outbound

This paper cites UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.994459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.994459Z digest=sha256:27e9c2ed7d3579db6598ba6b35e964d3804b4a14d5e4af8a5f6609f11f796193

Observation 7dfd02ad-5cab-496b-9e6d-33fe1fbad1e6 · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.044394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.044394Z digest=sha256:90995e1721238f88a15a78898450c745acd513e3360c886dd618769763ced374

Observation 9d390e1d-3709-4354-a1b9-ec425eae07b5 · outbound

This paper cites SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.109011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.109011Z digest=sha256:d767d3403fc55daa4bc676912455e85302a1f82aba351908997a33c53d356cd0

Observation d2aa8661-24bb-4b08-835d-a182292dc5d0 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Moshi: a speech-text foundation model for real-time dialogue

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.181884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.181884Z digest=sha256:16f34d7ebd9d8f7fd5dadd8e36ee44a7c3fe9da6d2ad8091ec74f8ec6d52e8ae

Observation bfb54c71-31ca-47cd-bcd4-8373d8efd4c3 · outbound

This paper cites Dualcodec: A low-frame-rate, semantically-enhanced neural audio codec for speech generation,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Dualcodec: A low-frame-rate, semantically-enhanced neural audio codec for speech generation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.245258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.245258Z digest=sha256:f5928e8581cddc2531075c9805a3a427dba8da96362f6d22924ac13209533b3d

Observation 48c2eb87-df7c-4182-adba-3d8e8ad3215a · outbound

This paper cites Brain plasticity under early auditory deprivation: evidence from congenital hearing-impaired people,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Brain plasticity under early auditory deprivation: evidence from congenital hearing-impaired people,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.315157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.315157Z digest=sha256:82f2377066bb6a75e6fb7b8ee7c3e0aad92182ddf52c871c4efa45d8e239bd75

Observation f3acf3a7-62b6-4c28-9b30-641c07a98d23 · outbound

This paper cites A tutorial on mpeg/audio compression,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners A tutorial on mpeg/audio compression,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.370793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.370793Z digest=sha256:6ac1d61feb8a566576b51c837bf368241bb262fbf3591015d350dd7701063d7d

Observation cf877525-a0bb-42e4-b250-0bafe52deed8 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Soundstream: An end-to-end neural audio codec,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.454439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.454439Z digest=sha256:078bbcabcbac37d30c554e2058ad0033573deba5c29bfe3afc06af0acf711964

Observation eccc40f4-61df-481b-af0b-307de4dfe04a · outbound

This paper cites HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.501429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.501429Z digest=sha256:51d642434d0d077ba78b8bfb4a13a6069d750c2c95e055b10d1ab9d813934007

Observation 2d2e3a6a-f07b-4d3d-893a-fa92edcd6f4e · outbound

This paper cites Flow-vae vc: end-to-end flow framework with contrastive loss for zero-shot voice conversion,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Flow-vae vc: end-to-end flow framework with contrastive loss for zero-shot voice conversion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.561461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.561461Z digest=sha256:ae8e25b482d8bef52a843f024703c909a3bc7eca372c82eb7d89a021a396c8d0

Observation bbc9e49c-ef69-47f3-9146-bd6ef4331b1a · outbound

This paper cites Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.614490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.614490Z digest=sha256:77b68e921156b45181ebdf99bbf11f68f2c8557bbed3943fff05cba47cc94b0e

Observation 5f859916-789d-445a-8fa7-feb35227b639 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.643574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.643574Z digest=sha256:9996556a9b25fb2177f267e18ea16a7f77998e67f5831d7a43e1f1cacaeac678

Observation fdc269f3-687d-46b1-bfab-0830b3832565 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Wavlm: Large-scale self-supervised pre-training for full stack speech processing,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.726141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.726141Z digest=sha256:0bf821dd15c13cb7a8cc7f808c64c6687fd5688fe684882960ee1b832b573a25

Observation 987f8729-1321-4bef-bda1-a959f73e3152 · outbound

This paper cites SUPERB-SG: Enhanced Speech processing Universal PERformance Benchmark for Semantic and Generative Capabilities.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners SUPERB-SG: Enhanced Speech processing Universal PERformance Benchmark for Semantic and Generative Capabilities

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.768658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.768658Z digest=sha256:1b3158a70a41cc85b14512ac3bbae2dacc78cec1dde37ba0eeefd929b49f573a

Observation 308b81e9-2564-4b84-92a7-c3107e254bc9 · outbound

This paper cites RepCodec: A Speech Representation Codec for Speech Tokenization.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners RepCodec: A Speech Representation Codec for Speech Tokenization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.824678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.824678Z digest=sha256:fb435a9a02370f22e455357242039b5cd368fcc44220f4cdc8935bf1de21a5a4

Observation c4672b52-6489-4b5d-9fc1-6ca23d2aeb53 · outbound

This paper cites Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.878282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.878282Z digest=sha256:1932eb695302978c8d2fad83e9609bc200115b5dccdbbd468b1876ea8b826652

Observation 5538d06b-21be-4960-9335-9ed3868c6660 · outbound

This paper cites Semanticodec: An ul- tra low bitrate semantic audio codec for general sound,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Semanticodec: An ul- tra low bitrate semantic audio codec for general sound,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.907602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.907602Z digest=sha256:cca353b34658aa039a861b4da1c2df2d9f217e7ef7710b5c67cbbf3eb6a31bcc

Observation fd8a5c3f-1518-47c4-a32f-e3ac2de11812 · outbound

This paper cites Sixty years of frequency-domain monaural speech en- hancement: From traditional to deep learning methods,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Sixty years of frequency-domain monaural speech en- hancement: From traditional to deep learning methods,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.960879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.960879Z digest=sha256:fe913894426e0769895cf29a07dc66fea708e26ff9c2f3bd227953db01a536ac

Observation 0c42ae0b-25e4-466e-a677-449103f0395c · outbound

This paper cites A review of vector quantization techniques,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners A review of vector quantization techniques,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.007585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.007585Z digest=sha256:e3bd2b8d788c7d3b11e17b2614ca87b55d59c1706f35b704a3b65b53cbe52b4d

Observation 3ac3e4db-175f-410b-80f3-72ee20d7fc32 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.066979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.066979Z digest=sha256:38c8dca8d53bdbf6f793b033e204a52a711d0a9c157d2da3eaca1b1084ca1d99

Observation c7d849c3-24ba-4204-89b7-9c77de90223e · outbound

This paper cites Neural dis- crete representation learning,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Neural dis- crete representation learning,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.096021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.096021Z digest=sha256:2bdbcfa89f733ddf84f5c3a108c2632d93cda7474a3dff32f90e85e4cdad6c70

Observation 3fe8a6d4-6c3c-4f6e-a61e-b88ac4cb3a0d · outbound

This paper cites AISHELL-3: A Multi-speaker Mandarin TTS Corpus and the Baselines.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners AISHELL-3: A Multi-speaker Mandarin TTS Corpus and the Baselines

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.149135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.149135Z digest=sha256:644aab71f9269f0732dd7b84924fb5093ab2d6cf273d5563404bfb02b327fac7

Observation 0ff7203d-9946-4890-96b3-d0be7e7a1bb0 · outbound

This paper cites LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.193305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.193305Z digest=sha256:9464681eb591475a8e158f84fc4b879ec748b810f987ce80e4a47761e283dec9

Observation b08ec887-b932-4c57-99aa-ded01908740d · outbound

This paper cites The voice bank corpus: Design, collection and data analysis of a large regional accent speech database,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The voice bank corpus: Design, collection and data analysis of a large regional accent speech database,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.249521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.249521Z digest=sha256:5bf6397914c29869363a6c1c3de90980767634479cad707f31f9c3db44d3b869

Observation 0c2628a5-a575-43e7-ad54-c9f4e6319597 · outbound

This paper cites The design for the wall street journal- based,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The design for the wall street journal- based,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.336254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.336254Z digest=sha256:c3f13734c6d21acc4fd93239a24b9dd66e88fa44721cdeb9b7e6b278b62cc721

Observation 2ade5bfd-d711-4079-85fb-10545fbdb348 · outbound

This paper cites Esc: Dataset for environmental sound classification,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Esc: Dataset for environmental sound classification,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.415756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.415756Z digest=sha256:71fc8bb917bfae3197cdef090ab54cb6c36e7ef8c9deac52b8cd4e2cf818bcc1

Observation ed606094-3758-4290-9f32-7b616663d3b6 · outbound

This paper cites The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.473571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.473571Z digest=sha256:9851fba30b5c228e671451c89f841e4dd0d7ecd08694fce82c3d7f0c6fe8b107

Observation 03832718-d907-4723-86a0-1af4f5511876 · outbound

This paper cites First stereo audio source separation evaluation campaign: data, algorithms and results,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners First stereo audio source separation evaluation campaign: data, algorithms and results,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.530768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.530768Z digest=sha256:ad970967ba85d3ebda7ee7ef3b08465cd1f2907ca9cf2c401a1a33659639d77f

Observation 9d4956ab-1ef7-4e02-b000-f4c7185a9688 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Robust speech recognition via large-scale weak supervision,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.584010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.584010Z digest=sha256:6b4c1971e466a85d382b73f8e82222a11f036ac6c2a5cc42c48f1a50251e7e61

Observation 7f27f1f1-edb8-4ca8-b669-ec4ca13cbc3d · outbound

This paper cites Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.644016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.644016Z digest=sha256:0b531d8900450e04354a0d02bc817fda4d7e1a799d758d6a16560acfa2b565e8

Observation 0df59a6a-045c-4892-b56a-b72933812da5 · outbound

This paper cites Inter-subnet: Speech enhancement with subband in- teraction,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Inter-subnet: Speech enhancement with subband in- teraction,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.669090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.669090Z digest=sha256:2bab028c1d206dbb233aa0495944db1af19544ffc25341b5ab043a68b3956deb

Observation ce48057c-768d-4a47-ac92-b8fc049ae857 · outbound

This paper cites Storm: A diffusion-based stochastic regeneration model for speech enhancement and dereverberation,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Storm: A diffusion-based stochastic regeneration model for speech enhancement and dereverberation,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.722918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.722918Z digest=sha256:12c1403f235a0349b4567bb126accbc82fe9b350b89d92a03894b74cb2d3d7d5

Observation 4c826304-4843-4d75-8b72-968118a0e324 · outbound

This paper cites Selm: Speech enhancement using discrete tokens and language mod- els,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Selm: Speech enhancement using discrete tokens and language mod- els,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.774367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.774367Z digest=sha256:bcdb94aea1d02a75d43a4ae988928f4ab46a6c89c659c121c8f767983b9217c3

Observation 5ca5a13c-4aa2-47f2-825c-8cc2979366aa · outbound

This paper cites Attention is all you need,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Attention is all you need,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.831514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.831514Z digest=sha256:70bcf5b0e6ce9f205b1b89227598b51ab48085eeae61ff78f56ddcea0263244b

Observation 34b8bdc9-f242-4121-ac11-2ed83fb76510 · outbound

This paper cites PolySpeech: Exploring Unified Multitask Speech Models for Competitiveness with Single-task Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners PolySpeech: Exploring Unified Multitask Speech Models for Competitiveness with Single-task Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.886688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.886688Z digest=sha256:a94cc7ebe09a18b190a873f95422d27850025db5a9a33c90a63f505b072e15fd

Observation f489a456-69c3-44f8-9163-9a4a535cf71a · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.944636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.944636Z digest=sha256:0ae8a1f5f4acf52e90acc16ebaf9cd102845f995b8909f2a08ff0423c200568f

Observation c6eee5b8-c5b0-4fd9-b5d5-aeff362fc939 · outbound

This paper cites Decoupled Weight Decay Regularization.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Decoupled Weight Decay Regularization

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.994131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.994131Z digest=sha256:c81ac774e21759475a9945d11a673d15a7601afcb38fd11c5db129749f370fec

Pith citing papers

Observation 418f8ce1-a402-44c1-90ce-25eb16f76ea9 · inbound

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners cites this paper.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.680596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.680596Z digest=sha256:9320332881fc36fcba66dcf96470b4faefce771259548bbac1319a6a432e5571