Pith. sign in

Paper Citation Record · LEDGER

HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2305.02765.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.02765 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 42 of 42 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:37:03.726917Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

19
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2db04592-9e87-449a-86a0-452f6c8ee9a9 · inbound

DASB - Discrete Audio and Speech Benchmark cites this paper.

DASB - Discrete Audio and Speech Benchmark HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:28:39.528794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T00:26:57.419537Z digest=sha256:af82e8b5dbaa29e41d87db3362b6dfe6e9f161c10a275296dea8902d1fc9decd

Observation e7ecd608-8cac-4b36-a2c6-bce8f1986ff7 · inbound

F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching cites this paper.

F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 151

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:06:41.558268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T06:06:41.348728Z digest=sha256:7b1757dd0c375779badfc9928d991fa90fcee23ed1204933474f0ea570b4249f

Observation 1322981d-a653-44ed-a58f-ded95262096f · inbound

GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling cites this paper.

GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T10:37:03.726917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:37:03.726917Z digest=sha256:b69931db966b87d8a688fabafff38bfb0bb9dd0bbdf8dbcda910313d7859b24b

Observation 8d564f31-603c-406e-b171-5f5efd47bbc6 · inbound

Do we really have to filter out random noise in pre-training data for language models? cites this paper.

Do we really have to filter out random noise in pre-training data for language models? HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-08T15:04:29.618212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:04:29.618212Z digest=sha256:1791a3beb8126d023be70d941c915ec3372f44fa2a8d3f6f96296f3bb4e4b1af

Observation ae2a569d-ba4b-42a9-a956-bcf254094187 · inbound

PAST: Phonetic-Acoustic Speech Tokenizer cites this paper.

PAST: Phonetic-Acoustic Speech Tokenizer HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:44.119792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:44.119792Z digest=sha256:e4aab43ff2b3855d8e4ded538dc2fe5b79350e647a1751de6863f63bfdb4c38b

Observation 253a6c96-6aff-4fbb-b24d-1f91f865a911 · inbound

EASY: Emotion-aware Speaker Anonymization via Factorized Distillation cites this paper.

EASY: Emotion-aware Speaker Anonymization via Factorized Distillation HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:30:36.404353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:30:36.404353Z digest=sha256:d721c01c8a8f8781de8e05178151fb7393de53a20b8830277ccc1f7f163a1fd1

Observation c17515ea-574e-4cd3-9cdd-f3aa64e2267c · inbound

UniTTS: An end-to-end TTS system without decoupling of acoustic and semantic information cites this paper.

UniTTS: An end-to-end TTS system without decoupling of acoustic and semantic information HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:53:17.856062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:53:17.856062Z digest=sha256:16fb5aea940144715649c6e6acd41ce652c9581e2cfa268543696a928cc5dad2

Observation cb3a9791-83df-42c5-b990-e4fc8f7d12cd · inbound

Vision-Integrated High-Quality Neural Speech Coding cites this paper.

Vision-Integrated High-Quality Neural Speech Coding HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:50:34.712516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:50:34.712516Z digest=sha256:89e6da3107d64c6de92e5fae54f5d6aa2bc6f4ffc5b15bc4131d0aa3bc8541a7

Observation 314333f8-cf24-4d99-ab6f-6c3aa6839524 · inbound

Probing the Robustness Properties of Neural Speech Codecs cites this paper.

Probing the Robustness Properties of Neural Speech Codecs HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:31:36.315885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:31:36.315885Z digest=sha256:820898d3ebd1ee3e2835aa732455858572de32e40a88c683c409e429f43abd46

Observation 25530ef0-39a0-44ea-99d3-c96a65232ecd · inbound

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec cites this paper.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.551229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.551229Z digest=sha256:5197cee34a047e83a7419c47afb6568dc1301bc6ffcb57f683d21832f53a65cf

Observation 7f80eabb-77c8-4a24-900b-cfcde254fa56 · inbound

SwitchCodec: A High-Fidelity Nerual Audio Codec With Sparse Quantization cites this paper.

SwitchCodec: A High-Fidelity Nerual Audio Codec With Sparse Quantization HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:12:18.263887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T13:10:14.839742Z digest=sha256:8c44cf378da76672161fe0fcb71e5e04e70617dd8aaddd0c0110bd2bc98ee43e

Observation 05bcfd94-028f-404c-86a9-ce88db9cb736 · inbound

Speech Token Prediction via Compressed-to-fine Language Modeling for Speech Generation cites this paper.

Speech Token Prediction via Compressed-to-fine Language Modeling for Speech Generation HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:30.136628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:30.136628Z digest=sha256:fe5d552421856c338b17ac9472fd2a54d146342bf263a175005415e38a8905e6

Observation 416732da-24e4-4e09-9e70-d2114e6776df · inbound

Collecting, Curating, and Annotating Good Quality Speech deepfake dataset for Famous Figures: Process and Challenges cites this paper.

Collecting, Curating, and Annotating Good Quality Speech deepfake dataset for Famous Figures: Process and Challenges HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:23:52.752336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:23:52.752336Z digest=sha256:d9c5b748ac5d2c63b515df64f0c9c59034a087bfa56158a5315c81d07dbaae1c

Observation c46c1d72-99e2-425d-b195-235d3ac823e7 · inbound

Autoregressive Speech Enhancement via Acoustic Tokens cites this paper.

Autoregressive Speech Enhancement via Acoustic Tokens HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:41:22.610257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:41:22.610257Z digest=sha256:a92523aacfe848b816c67062b19d26b285dfa19b19486c6c77dab6420c5f6631

Observation acbb01f9-ab9e-405a-9aaf-dc55f47d8d26 · inbound

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos cites this paper.

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:33:40.225699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:33:40.225699Z digest=sha256:a028f756633c34ca49e84c1aebec91e6f58b798958f5c9f041fe15fbe8eb16ba

Observation f225f353-cfd7-4e63-b6bc-70213e002bb5 · inbound

Representing Speech Through Autoregressive Prediction of Cochlear Tokens cites this paper.

Representing Speech Through Autoregressive Prediction of Cochlear Tokens HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T19:54:58.565677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:54:58.565677Z digest=sha256:9c5fdd0ecbde68e70ee1a3899819a48a9b0d5263adf5f68944bd68d72917173d

Observation 06b61e99-6fde-4deb-a33f-daeb94aad832 · inbound

MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations cites this paper.

MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T15:58:54.074017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:58:54.074017Z digest=sha256:21c635f94e0a78da81079576fae1f125c718c1815d48757dfdc669cab53fe52b

Observation 83bbd839-6360-4737-9f50-a666710a8e05 · inbound

Analysing the Language of Neural Audio Codecs cites this paper.

Analysing the Language of Neural Audio Codecs HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T12:38:50.189716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:38:50.189716Z digest=sha256:d935ea7de6ad856e67a636d14d9f645fb23b2fee7df7b6448615bd2f9a80303b

Observation 9e4cf411-10da-45fd-ad7f-82ba079270fc · inbound

AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation cites this paper.

AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T11:41:03.195908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:41:03.195908Z digest=sha256:31260a275961dd627d1bd8e3d13e08b99720e7c25f3af5cc9e1bf8709261c08f

Observation eccc40f4-61df-481b-af0b-307de4dfe04a · inbound

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners cites this paper.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.501429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.501429Z digest=sha256:dbf7a63b4b74f855c2ef1457ae66acb1c1b00e658d4c7a7a3a32d997a1f8afb4

Observation 81a5ce7d-c10b-49f3-9b0e-09a86d317a0e · inbound

Two-Dimensional Quantization for Geometry-Aware Audio Coding cites this paper.

Two-Dimensional Quantization for Geometry-Aware Audio Coding HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:20:29.324852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T18:16:51.486807Z digest=sha256:c2c70e2e32c015e97d9b912ea9dd186da46624b4287fcd860a7d697ca0cd49ca

Observation 8765431d-586e-40ac-ba53-632d9fa2ae65 · inbound

The Equalizer: Introducing Shape-Gain Decomposition in Neural Audio Codecs cites this paper.

The Equalizer: Introducing Shape-Gain Decomposition in Neural Audio Codecs HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T22:52:22.957365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:52:22.957365Z digest=sha256:29e4d6ddc615ef9b9343c29a6981a4a2e4f4d542980b823cf0351e8610fdc27a

Observation 13722c82-0799-4806-a3a0-cbcc3866a4ab · inbound

On the Distillation Loss Functions of Speech VAE for Unified Reconstruction, Understanding, and Generation cites this paper.

On the Distillation Loss Functions of Speech VAE for Unified Reconstruction, Understanding, and Generation HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:25:59.517768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:30:58.664375Z digest=sha256:d78901ae811f1be018f5daadc92fb8bf3686443aeb5f6272260431edea0b5a86

Observation 306e6cdf-ba9c-42ec-bb99-acc2fcc58841 · inbound

Beyond Static Collision Handling: Adaptive Semantic ID Learning for Multimodal Recommendation at Industrial Scale cites this paper.

Beyond Static Collision Handling: Adaptive Semantic ID Learning for Multimodal Recommendation at Industrial Scale HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-08T23:39:24.351337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T05:41:17.427618Z digest=sha256:3fd941f7c459316208ed4f7fd17488b37a25b77eaa84e42cef9d2045a51e547a

Observation 7662c559-d436-46f3-9ae2-6c42e0f6c8c7 · inbound

SPG-Codec: Exploring the Role and Boundaries of Semantic Priors in Ultra-Low-Bitrate Neural Speech Coding cites this paper.

SPG-Codec: Exploring the Role and Boundaries of Semantic Priors in Ultra-Low-Bitrate Neural Speech Coding HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:11:25.429100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T12:44:08.364835Z digest=sha256:c667a3dbd9863f29e7c8432401dee3bf8fc2dec41627365769d3b547677705a4

Observation 548a9c9b-923b-4f31-83a7-bbd54a3fb979 · inbound

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization cites this paper.

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:08.827723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T12:25:52.847432Z digest=sha256:b3d2cb350244038c70211adeec6399b79e54f792fbf944356b2ee91e5a1c2233

Observation 55b89657-e9ed-471e-98c0-a7d228024d93 · inbound

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization cites this paper.

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:15:07.755327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T23:14:32.494076Z digest=sha256:a7b934f9915d43dfa769f9a4f9b7314a5cc9d58309df891a7c5e5d9493b82c68

Observation 07ed2eae-dc0b-438d-af05-df6a276367da · inbound

AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling cites this paper.

AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:07:00.268282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T01:04:54.506749Z digest=sha256:74c56cf13865e929c384946406606bc54427865db9c2eb3e101c0e98fa399464

Observation e00e85cb-1fc9-47d9-9ffb-178713e42b61 · inbound

Optimising Neural Speech Codecs for 300bps Communication using Reinforcement Learning cites this paper.

Optimising Neural Speech Codecs for 300bps Communication using Reinforcement Learning HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-20T02:07:58.601955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T02:07:50.016394Z digest=sha256:616bf76e55a0c66725aa3423506456a12ed86959873f60bf91be67679b647aed

Observation 32f8e5fd-7e83-4c60-a83b-fcd812423b6d · inbound

Optimising Neural Speech Codecs for 300bps Communication using Reinforcement Learning cites this paper.

Optimising Neural Speech Codecs for 300bps Communication using Reinforcement Learning HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T13:42:35.682547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:42:35.682547Z digest=sha256:f5b0075ab98925d681738fceb471217b7fdee1c3a7f181b7a766d6bd692f5c03

Observation 66bb1792-8c77-4365-bc42-ac9dbf53472a · inbound

AffectCodec: Emotion-Preserving Neural Speech Codec with Block-Diagonal Residual FSQ cites this paper.

AffectCodec: Emotion-Preserving Neural Speech Codec with Block-Diagonal Residual FSQ HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-25T03:05:16.739256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T03:04:01.544231Z digest=sha256:8bcc07face538a43d6a4e63bd0ec3897d151551f39ebf8a8454fa7def8e5ecbc

Observation cb4ffe68-c56e-4813-9e8f-8f240219b418 · inbound

UniAudio-Token: Empowering Semantic Speech Tokenizers with General Audio Perception cites this paper.

UniAudio-Token: Empowering Semantic Speech Tokenizers with General Audio Perception HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 6

Resolution
malformed identifier
arxiv_id, observed 2026-06-28T22:32:44.571649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T22:26:20.100595Z digest=sha256:3fdc94287b3c1f70684af3c75cb19fd659e3d0f9a8a4c26a1f1157580842e9b7

Observation 547f23bc-24b8-4173-acaf-1a4af0ad7839 · inbound

ContextCodec: Content-Focused Context Guidance for Ultra-Low Bitrate Speech Coding cites this paper.

ContextCodec: Content-Focused Context Guidance for Ultra-Low Bitrate Speech Coding HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:47:44.553305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T11:51:25.305324Z digest=sha256:3ca2fd6e4058bfe9a72d86bf373d9babdba6647befbbfe94265bc881a45ac32e

Observation ca0eba20-563a-4f04-a5ee-e13b8ed8c59f · inbound

Benchmarking Neural Speech Compression from a Rate-Distortion Perspective cites this paper.

Benchmarking Neural Speech Compression from a Rate-Distortion Perspective HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T12:48:12.256905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T08:43:35.279033Z digest=sha256:9b3f320080aa9d7165bbdc29249936054ff91b58f9268428706ef6ff0aa7ffb0

Observation 2f821992-797c-421f-825a-db2758e53f04 · inbound

Self-Guidance: Enhancing Neural Codecs via Decoder Manifold Alignment cites this paper.

Self-Guidance: Enhancing Neural Codecs via Decoder Manifold Alignment HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:08:37.514482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T06:05:26.735340Z digest=sha256:904fe9c28a73fa9eac28ebea634d7a4ba7bff4e15b1d168ef2f8c75ccba64bcb

Observation a64f6781-5bea-4433-95d9-0482ae2e5f24 · inbound

Self-Guidance: Enhancing Neural Codecs via Decoder Manifold Alignment cites this paper.

Self-Guidance: Enhancing Neural Codecs via Decoder Manifold Alignment HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-06-28T16:02:22.518914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T06:05:26.735340Z digest=sha256:145ffa65c78baa69af6722ee052deeee58226f34e318c2208c58008401d5b341

Observation 781fbedf-bd44-4908-bfd8-3ad9a33c9f75 · inbound

FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model cites this paper.

FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 122

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T11:55:42.208918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T03:50:26.873406Z digest=sha256:2c0339ee68903841cb21b007a0ead58bd38c0c0ab91dbc15bef51833aeaf82c1

Observation bef5c968-b30d-4668-9998-c1451e4820ea · inbound

Positive-Incentive Noise Predictor for Adversarial Purification in Speaker Verification cites this paper.

Positive-Incentive Noise Predictor for Adversarial Purification in Speaker Verification HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-02T05:16:38.675904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T05:12:44.574336Z digest=sha256:efe25d57503e2aa4df597e1f27b5bf37e933597ca675ea7edbec4ea7c284a241

Observation c29c4570-5299-4921-8488-10667aa4eef9 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 245

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.350821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:d0a30993c895aabd765d8a2605d56069a264bf37b55bb385276b7b76e1229556

Observation 9b6c5251-c268-46da-86cc-432ca6c6da52 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 245

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:1b142f73ccd0156fbb6bb933609c0599a90d658e5be2b5c8de6dbe5dd802da6a

Observation f7debbb3-ad19-48b2-9569-06106a7b6899 · inbound

ReGen: Hierarchical Multi-Prompt Representation Generation for Efficient Waveform Diffusion Models cites this paper.

ReGen: Hierarchical Multi-Prompt Representation Generation for Efficient Waveform Diffusion Models HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-13T05:10:26.667731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T05:10:26.667731Z digest=sha256:122f47e924a627ec8e0263a0a2f21f96e32c0b5a8aa164d8a439d24e57a03af2

Observation ab5e6bd2-fdfb-4d2d-8562-9a301254364b · inbound

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness cites this paper.

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T08:25:27.374503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:25:27.374503Z digest=sha256:20c8d872b33ee9209b016000cd7f056c5331858557bc6c0666728dc07fe623f2