Pith. sign in

Paper Citation Record · LEDGER

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec

As of 13 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 2 inbound Pith citation observations for arXiv:2505.24314.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24314 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:42.730075Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:38.143308Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T12:58:08.378519Z

Reference resolution

36 of 36 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2fa164f2-80ec-4a8d-b819-b94dfee2e5c0 · outbound

This paper cites A pivotal challenge is the transformation of continuous speech signals into interpretable representations suitable for inference and training within large language models.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec A pivotal challenge is the transformation of continuous speech signals into interpretable representations suitable for inference and training within large language models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:45.426994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:38.015315Z digest=sha256:6c296c22869d1729e2ede98c3a427910791a7f8a76e0e1178d9fb43ae628205b

Observation 76b3dc45-1f91-4739-b897-3356dcbba6c3 · outbound

This paper cites DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.143308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.143308Z digest=sha256:cdb60543b05c1336358f008ac4f8e04b1962b0daf16ea1230182b249c391173f

Observation c0c0a5dc-6434-47c6-b322-762abc0e66cd · outbound

This paper cites Dataset and Metrics We use LibriSpeech [27] to train the speech codec we proposed.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Dataset and Metrics We use LibriSpeech [27] to train the speech codec we proposed

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:45.216566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:38.356758Z digest=sha256:01a40e6e4811b0314dcdfb2aee591cc3164bc240d38148e9c5832e80fafc9d7f

Observation 50daa269-e8bb-427e-85fa-e338e3a0802c · outbound

This paper cites an unresolved cited work.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:29:45.003225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:38.550836Z digest=sha256:61f6535bec25ffb636bef83d052606ed24590d528444e8a8fed0fd2dac4eb9f8

Observation 135c561d-84a1-4b24-ad73-6fd213a3eb8b · outbound

This paper cites an unresolved cited work.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.724956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.724956Z digest=sha256:9a9bfa2216646d12e0573443f9058a92c0fc63159ab5304af6fb432c3fd8fb36

Observation c93301bc-58d4-4e4a-895d-74c5133d163b · outbound

This paper cites GPT-4 Technical Report.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec GPT-4 Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.877708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.877708Z digest=sha256:e9fcf588ae3172beeff5890dad9e30653bf8d8da53cebaf447d9f7cc72596e01

Observation df8710a3-de5f-4dc5-a521-9e6086b4050e · outbound

This paper cites GPT-4o System Card.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec GPT-4o System Card

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.995350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.995350Z digest=sha256:680462c55b42fc86a5cf9553de4f42fe2367d94539bac47eb81ce84dc969d829

Observation 66668e64-fd4e-4e02-80ef-9a0e7aedb0c6 · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.170430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.170430Z digest=sha256:1ebb159494af6069ec908bbcc325e72b90090c74c80d450f68917c429238cbc4

Observation 4d20e720-c46f-429d-940e-18ebc9537d45 · outbound

This paper cites Audiolm: a language modeling approach to audio gener- ation,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Audiolm: a language modeling approach to audio gener- ation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:44.799493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:39.311348Z digest=sha256:c3955ff0e2837da3c0704f48899fcf89759076e079c5b534fc9558e06f018bb0

Observation dd6f89eb-99a1-4a4f-b2c3-2541d03efd3c · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec AudioGen: Textually Guided Audio Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.487610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.487610Z digest=sha256:0926fefe8cda319f3dd3196b19cbb771e2870f66996a8a56169f150a2019e4d4

Observation 52a05e72-d8c6-460b-8948-185223fca620 · outbound

This paper cites Slim- speech: Lightweight and efficient text-to-speech with slim recti- fied flow,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Slim- speech: Lightweight and efficient text-to-speech with slim recti- fied flow,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:44.553668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:39.605523Z digest=sha256:57fee407a9ebaf249ace278ba3efd14cedf997b9eef9988b52d81f14385558ae

Observation 8aed8f60-4168-402d-833e-330fb7cca6b2 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.703511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.703511Z digest=sha256:379e75506f3166da562ea28fabe8961af8c9e7d1df3ad43daadede0ab2ba1ed0

Observation e7f091b9-a0e6-4a24-993f-f5a1d4d7890b · outbound

This paper cites wav2vec: Unsupervised Pre-training for Speech Recognition.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec wav2vec: Unsupervised Pre-training for Speech Recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.796888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.796888Z digest=sha256:049e6e8a76fe17d5b3cad357cfdb95da5cf52fdfa4ed3f7f6099d233fa5f9388

Observation 7fbe07b5-6546-48b6-b90f-3772504bdf95 · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.892248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.892248Z digest=sha256:dc2612084d9580016d6df699ea5356b2d515116647e711703e0abb2a3cbd53f3

Observation 78a5e2e8-1abb-47ea-85e7-721a5bb96f9d · outbound

This paper cites Neural discrete represen- tation learning,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural discrete represen- tation learning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:44.307824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:39.994883Z digest=sha256:8ca36c2b16ba49eb8820a00e95d36443e7e6c612c68c775319fc46c73a663d76

Observation 2523b863-a545-4f77-a382-86958b415664 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Soundstream: An end-to-end neural audio codec,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.070603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.070603Z digest=sha256:85ef628b8e3340f17e8c6dd78f1bfaf58612956048c96da2caae33d6d4f1495e

Observation c6b6e291-50a2-419b-bda5-1d403f010a1e · outbound

This paper cites A review of vector quantization tech- niques,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec A review of vector quantization tech- niques,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.192413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.192413Z digest=sha256:a4cb720bcb5d66752bb40618a71653ed5275a3bfbb3e0a44fa8475a7a308a55c

Observation 4f05c427-1b02-471e-9010-8995781d9162 · outbound

This paper cites High-fidelity audio compression with improved rvqgan,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec High-fidelity audio compression with improved rvqgan,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.359010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.359010Z digest=sha256:f5b246810fde9a1ac32fb01eb5eb1985f6258dcf38a34618588ee4daa7163da6

Observation e7be356b-fd57-472e-a7d5-42ef4f47fad2 · outbound

This paper cites High Fidelity Neural Audio Compression.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec High Fidelity Neural Audio Compression

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.426238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.426238Z digest=sha256:4e4d656c03edfa9866efa08addcaad009076d2f4120de1ebf54a0242c2d79e8b

Observation 25530ef0-39a0-44ea-99d3-c96a65232ecd · outbound

This paper cites HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.551229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.551229Z digest=sha256:e0d619a26d895c3f619dbc57b1024e5f3ef28df5af82aca6a9d36a6020541df8

Observation be0e75f5-0f79-433d-a3df-d9d60d153304 · outbound

This paper cites WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.695004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.695004Z digest=sha256:b16e9b6464b41db6fe5612a837a2b7fda1f61dfb162870421f6ea9a35b7472c6

Observation 18ce0122-83d8-410c-a52c-5e72d928fc6a · outbound

This paper cites BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.791796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.791796Z digest=sha256:014aa299fce2a059cb4330ae2253869ad94bc54267358dc5e95ae5b54a8d6378

Observation dd699265-cf72-4969-8c57-dfa127b10c6c · outbound

This paper cites Single-Codec: Single-Codebook Speech Codec towards High-Performance Speech Generation.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Single-Codec: Single-Codebook Speech Codec towards High-Performance Speech Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.894154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.894154Z digest=sha256:c53de8ed66ad2efbd80f7e1d8e0f1d15a4fe4ce684be6bf990386d3f65eea4d1

Observation 9bd0c454-47d4-4e4c-a0f1-a5ed8524626a · outbound

This paper cites Addressing Index Collapse of Large-Codebook Speech Tokenizer with Dual-Decoding Product-Quantized Variational Auto-Encoder.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Addressing Index Collapse of Large-Codebook Speech Tokenizer with Dual-Decoding Product-Quantized Variational Auto-Encoder

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:29:43.140544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:41.008351Z digest=sha256:e9983c39dc8191fa4dcdd6da5fea38948b014c4b570c90c4bb0ad0dd32aa1dc4

Observation 446a5104-842b-467a-b0a4-d240d0b57b5b · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.168284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.168284Z digest=sha256:518130bf0326579ef4b1d6c70b6cc2692cfd92987e4999fb5345614ef7948180

Observation e4799653-6b8b-4791-8dde-31578b37c5f6 · outbound

This paper cites Neural networks fail to learn periodic functions and how to fix it,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural networks fail to learn periodic functions and how to fix it,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.272611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.272611Z digest=sha256:ad123b964b8d5655a511f6bb7db7c626773cfe09cf29c0d7d17883001f37aeef

Observation 88b08a74-0353-4fc9-bb31-d1c8306b6e77 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec LLaMA: Open and Efficient Foundation Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.423921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.423921Z digest=sha256:a44a83243e3419c9a21288cd8aaa14cd90a5e0e67639db45e3acbcc8e01bc821

Observation ea8bc6fa-13d4-4dc6-b9ed-2dc502f00316 · outbound

This paper cites Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.569951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.569951Z digest=sha256:49442e0499a5918778009de28e2e68ec2535dfffe153200959b9d3d5377fc0ba

Observation 72a21425-ef2f-45e1-b045-18c894474aed · outbound

This paper cites Vector-quantized Image Modeling with Improved VQGAN.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Vector-quantized Image Modeling with Improved VQGAN

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.764813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.764813Z digest=sha256:50a31a9fd6c0ebe5843a75d7a20d61dc240d575be6683fe9a11cc0f44d405abc

Observation d7dfcef2-cb64-443f-964a-127c625c7164 · outbound

This paper cites Product quantization for nearest neighbor search,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Product quantization for nearest neighbor search,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.908267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.908267Z digest=sha256:5bf46d663d8a27b08c2c9a974f8e5603fb93dc0b59359121e527bc6a064ad435

Observation 4d360477-8cb8-4fd3-8c03-57363716f9aa · outbound

This paper cites Apcodec+: A spectrum-coding-based high-fidelity and high-compression-rate neural audio codec with staged training paradigm,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Apcodec+: A spectrum-coding-based high-fidelity and high-compression-rate neural audio codec with staged training paradigm,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:44.001741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:42.045168Z digest=sha256:6f12b664cb5769f7ea0c1dadcc73c7c48527146c86775942c1df5212423d34b6

Observation 78882b71-3869-46ee-be91-9f552fe4d393 · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Lib- rispeech: an asr corpus based on public domain audio books,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:42.189677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:42.189677Z digest=sha256:96fa8385430eef839cf71255a33f418c015424255202048a1ecde1648207f8a0

Observation f7213156-5f1e-49bc-937f-5e62869b6ad6 · outbound

This paper cites Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:42.327056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:42.327056Z digest=sha256:56d453566acac673bcbde6141c00ed671abd4bbfba8dc2f7a312bdf4642232e2

Observation e454bdb2-bee3-4c12-b1ac-760223fe8ed8 · outbound

This paper cites UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:42.472159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:42.472159Z digest=sha256:dbcccdd36cfe77a81ba8f23cd70978cdbcb2cfe537826324273aeddafa51cd90

Observation 59040c31-6b56-4635-ba6f-3724689adf01 · outbound

This paper cites Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:42.622160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:42.622160Z digest=sha256:ce91710ced586472d54625e197419f82cad0a803c832efd72b5a91dc2a0f59a0

Observation 83dbc189-7a8e-4ccf-b0a1-1a46cc871e63 · outbound

This paper cites Fewer-token neural speech codec with time-invariant codes,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Fewer-token neural speech codec with time-invariant codes,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:43.717631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:29:42.730075Z digest=sha256:de7b80b2fc24e1da355cf107b2ad546d23e570583390ad123480bebb259df9c7

Pith citing papers

Observation 76b3dc45-1f91-4739-b897-3356dcbba6c3 · inbound

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec cites this paper.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.143308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.143308Z digest=sha256:cdb60543b05c1336358f008ac4f8e04b1962b0daf16ea1230182b249c391173f

Observation 8fceee68-b136-429c-92a9-30fb9d09d7bd · inbound

SARA: A Dual-Stream VAE for High-Fidelity Speech Generation via Integrating Semantic and Acoustic Representations cites this paper.

SARA: A Dual-Stream VAE for High-Fidelity Speech Generation via Integrating Semantic and Acoustic Representations DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T12:58:08.380228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T08:38:51.371500Z digest=sha256:12824d8f764a4efd0d63f879a5da4f242a9788a6e7caf7cfeb55671d6efb1268